Sakana AI Brings Recursive Self-Improvement to 'Physical AI' for Robots That Learn on Their Own
Japanese AI startup Sakana AI says it is expanding its RSI Lab to pursue world-model-based physical AI research
AI models, services and robotics
Japanese AI startup Sakana AI says it is expanding its RSI Lab to pursue world-model-based physical AI research
AI influencer runs it offline, finds it understands existing code and even generates new examples
Muse Glimmer is an agent-specialized model that runs on 24GB-class consumer GPUs with 4-bit quantization
Caching teacher model logits and a memory-efficient KL loss enable long-context distillation on a single GPU
Anthropic to enable auto-approval mode by default for Pro, Max, and Team accounts starting August 14
Generation halts past 100K tokens when run locally with OpenCode; resuming with a command restores work
MoE architecture with 37B active parameters, FP8 quantized version uploaded to Hugging Face
Adding two GPUs to 128GB of DDR4 yields 176GB total, with 2–3 tokens per second depending on quantization
Multimodal model handling images, video, and text released open-source on Hugging Face
New version adds editing tools and templates, Elo score trails OpenAI's model by a narrow margin
AWS unveiled a native search feature that manages operational data and embeddings together without requiring a separate vector database
Server-side built-in tool grounds model responses in up-to-date web knowledge, activated with a single parameter
Sessions lasting up to 14 days with GPU support expand production AI agent infrastructure
Fully open under OpenMDW license, tops LingoQA benchmark
Research preview argues harness design determines benchmark performance, released as open source
31B-parameter vision-language model interprets qubit diagnostic results to automatically generate calibration values
A workflow that finds top-level catalog assets via natural language and auto-generates datasets and topics
Combining KAI Scheduler and vCluster lets multiple teams use a single GPU as if it were their own independent cluster
Amazon Bedrock has adjusted on-demand inference pricing for its OpenAI GPT-5.6 model lineup
New policy refinement engine diagnoses failed tests and proposes formal-logic fixes
Formula One built a Data Accelerator with AWS to automate operations on its marketing data platform
Open-source project built on Gemma 4 and LiteRT-LM, with a frontend for small displays
Alpamayo 2 Super unifies trajectory generation, reasoning, and auto-labeling into a single model
Cosmos 3-based World Action Model suggested as a replacement for VLA robot policy architecture