Anthropic has Claude tackle AI alignment research, and it outperforms humans
Claude fixed 10 types of alignment failures on its own and beat 28 human safety researchers, while a monitoring system caught 39 attempted cheats along the way.
AI models, services and robotics
Claude fixed 10 types of alignment failures on its own and beat 28 human safety researchers, while a monitoring system caught 39 attempted cheats along the way.
Tencent says the model beat GPT-5.6 on a coding evaluation and even helped optimize its own training and inference process
A Labs prototype bundling paper tracking, code conversion, and hypothesis generation
In an interview video released by Google DeepMind, the robot Apollo said tying a trash bag was harder than balancing during sports
Anthropic is pulling its temporary 50% boost and replacing it with a permanent 25% increase
Anthropic rolled out updates to Claude Code this week, speeding up CLI startup, adding token usage visibility, and making auto-mode rules editable
Standard seats are free, while premium seats with 5x usage cost $15 a month for a year
OpenAI released a video showing how ChatGPT Work plays Spotify tracks and adds calendar events by switching back and forth across desktop programs
Labs verified by a principal investigator get a year of free standard seats, plus premium seats with five times the usage cap for $15 a month. Anthropic is also expanding its AI for Science credits beyond biology to every scientific field, offering up to $50,000 per project.
A VP of Research at Google DeepMind discussed how to handle uncertainty, from weather forecasting to robotics
A common interface that cuts equipment integration from weeks or months to hours or minutes, starting as a research preview
Anthropic released a cookbook linking its agents to Vercel's Chat SDK
An encryption method that keeps both model weights and test questions hidden from each other cuts the risk of benchmark manipulation
In an interview, Thibault lays out pricing and agent UX strategy
Claude has Chat and Cowork, ChatGPT has Chat and Work — this is turning into an industry-wide problem for AI apps
Having missed its window in coding, OpenAI is winding down Sora and Atlas to double down on Codex, agents, and Astra. At the same time, a sandbox-escape incident has frozen its largest training run.
Prime Intellect says the gain came from changing the execution shell around the model, not the model itself. The code is on GitHub.
Perplexity says weaving sessions, files, and sources into a knowledge wiki lifted accuracy by nearly 9 points while cutting token usage 15%
The 125B MoE model runs on RAM alone with no GPU required, and it beat Claude Opus 4.6 Max across several benchmarks
When US commercial AI models refused to analyze the attack logs, Hugging Face installed Zhipu AI's open-source model GLM 5.2 directly on its own servers to investigate the breach. It's the defender's side of the story that OpenAI's report never covered.
Educator count tops 300,000, with a shared data privacy agreement now covering 16 states
US usage hits 460 million messages a week during the school year, staying above 180 million even in summer
The foundation model trains only on continuous glucose monitoring data and predicts diabetes risk more accurately than existing models
Google DeepMind revealed a custom vocabulary screen, pointing toward better recognition accuracy for specific terms