MiniMax announces agent plugin support for MiniMax Code
The MiniMax_Agent account posted on X about agent plugin support in its coding tool MiniMax Code
AI models, services and robotics
The MiniMax_Agent account posted on X about agent plugin support in its coding tool MiniMax Code
Alibaba's Qwen team disclosed its own model's benchmark rankings
New feature controls per-user traffic by request, token, and connection units
Scores 38 on Artificial Analysis Intelligence Index, sits on the Pareto frontier for performance per parameter
Reference inputs expand beyond text, image, audio, and video to include documents and spreadsheets
SDK v3 automates generative AI inference optimization, from benchmarking to deployment, directly in notebooks
Patch release upgrades the grader model and incorporates the latest τ³-Banking version
Only an introductory video has appeared on the official YouTube channel; details remain unconfirmed
Preview version measures the ability to distinguish between when to help a student and when to let them think for themselves
Anthropic says it adjusted safeguards to reduce blocking of health and education questions
Company says it was rebuilt on the open-source Pi Agent framework, but detailed specs remain undisclosed
Following an open-weight letter signed with more than 200 companies, NVIDIA unveils an open model for robotics and autonomous driving
Non-developers can generate multi-tenant web apps in seconds by describing requests in plain language
AWS connects Codex metrics to Amazon CloudWatch via an OpenTelemetry collector to help teams understand usage and cost by group
Six skills let coding agents handle the full process from policy drafting to deployment
Following up on its Chess Puzzles benchmark, Epoch AI measures reasoning ability with a puzzle set from an undisclosed game
DeepSeek-first cascade solves more tasks than Luna alone on DeepSWE benchmark
Temporal policies built on new open-source policy language Dogwood, plus gateway rate limiting, aim to block cumulative risk
Combines Agent Skills packaging and MCP server configuration into a single reusable format across agent clients
Rollout targets Pro, Max, and Team users; separate classifier catches 89% of risky commands
An RLM harness for coding and long-running autonomous tasks, built around a structure that modifies its own harness state
Alibaba's new model scores 56 on Intelligence Index at $1.14 per task, but Kimi K3 wins on cost efficiency
Vercel Container Registry added the ability to make repositories public on August 7, 2026. Previously, read access could only be shared with up to 100 teams, but now any team with a Vercel account can pull images.
OpenAI disclosed on August 7, 2026 that internal evaluations of its upcoming model, Astra, could not rule out a "Critical" cyber capability rating under its Preparedness Framework. The prior model, GPT-5.6-Sol, had been rated "High" on the same scale.