Local LLM Community Flags a Gap in New 8B-12B Models
Reddit user points out lack of new mid-sized models for 16GB MacBooks, sparking debate
AI models, services and robotics
Reddit user points out lack of new mid-sized models for 16GB MacBooks, sparking debate
Apple Research presents a method to create preference data without extra annotation or external models
0.8B and 2B parameter models beat larger models on French-language benchmarks
Not video, not Gaussian splatting — an approach that assembles worlds from editable game-ready 3D assets
Stony Brook University researchers used Ai2's infini-gram to quantify overlap between AI-generated novels and existing literature
Microsoft has released a guide showing how to build work agents using only natural language in Microsoft 365 Copilot's 'Agents' menu
Apple found that unusually large-value tokens forming inside image-generation AI degrade image quality, and proposed a fix
"DLR-Lock" research lets users run open-weight models as-is while blocking modification
Analysis finds diffusion language models lag in long-context tasks while autoregressive models pull ahead in batch processing
Fully open AI lab Ai2 increases its Hugging Face Hub storage to 2 petabytes and removes download speed limits
Trained on 2.1 trillion tokens, then self-distilled to generate sentences in as few as 4 steps, testing an alternative to autoregression
DYNA-2, a robot foundation model pretrained on 170 years' worth of human behavior video, emerges
Zhipu AI has reset usage limits for all GLM Coding Plan users and rolled out an update to strengthen long-horizon task handling
Workbench Notebooks extension launches, directly linking local IDEs to Google Cloud compute
Upgrade to open-source character animation model includes multi-character support and text-driven camera control
MIT-licensed CLI and gateway store work history and past failures in a .memcode folder
Artificial Analysis to hold event in San Francisco on August 12 addressing speed gaps across inference providers
Aidan Gomez explains why the company is expanding its open-source lineup and how it's designing for "sovereign" deployment
The independent evaluator known for measuring model performance is now preparing a tool developers can use directly
Nous Research adopts scripting approach for browser agent, cutting token usage by up to 66%
Tiny corp released a video showing a 7900XTX connected over USB3 running Qwen 3.6 27B. The dock ships on the 12th
Security firm PromptArmor discloses vulnerability letting Rovo exfiltrate Jira and Confluence data without a trace
Company says it couldn't solve the hypothesis itself but raised the lower bound on the proportion of zeta function zeros
Anthropic switches its coding agent Claude Code to auto-execution mode by default