GLM-5.3 API released, Terminal-Bench score jumps from 4.6 to 28.3
Z.ai opens GLM-5.3 API at the same price as GLM-5.2, with sharp gains on coding and cybersecurity benchmarks
AI models, services and robotics
Z.ai opens GLM-5.3 API at the same price as GLM-5.2, with sharp gains on coding and cybersecurity benchmarks
Connect via the connectors menu for email drafting through sending, with users setting when approval is required
Describe a reverb or delay in plain language, and Suno Studio generates the plugin in seconds
Anthropic had external labs validate protein binders that Claude autonomously designed for 14 of 15 target proteins
Three OpenAI employees walk through their step-by-step workflows for building meeting briefs, blog drafts, and strategy decks with ChatGPT Work, alongside the full text of prompts OpenAI released
The beta that started with the Max plan in July has widened to all paid plans in six weeks
Artificial Analysis released a benchmark showing agent performance jumps from 33 points without search to as high as 75 with a search API
Add computer@perplexity.com to a thread and the email conversation becomes an AI work session
A screen-control feature already available in Claude and ChatGPT has been spotted inside Google Gemini app settings
Partnership with Exa brings web-search citations, plus beta expansion into automatic tab organization and natural-language history search
From study mode and homework reminders to five parental controls — details confirmed via OpenAI's official guidance page
JAMA authors, including a University of Pennsylvania bioethicist, call for regulatory reform in preparation for an era when autonomous AI outperforms doctor-AI collaboration
Penn State study finds instruction loss during context compaction; small add-on module restores retention to over 90%
Alibaba's Qwen highlighted the competitiveness of its 27B model by promoting a user's local-run comparison case
Anthropic analyzed over 300,000 conversations and found that Claude expresses different values depending on the model and language
Six months with the $899 AI robot litter box — waste removal was excellent, but facial recognition and waste detection fell short
Open-weight MiniMax-Music3 generates a complete 32kHz stereo track in one pass from just two inputs: lyrics and a structured caption
Anthropic has added an artboard-based design feature to Claude Code as a research preview
Each bot gets its own role, model, memory, skills, and profile picture — and bots can even talk to each other
Hypha AI and Rogo AI cases show token-saving strategies, Sol triples ARC-AGI-3 score
Base44 shares performance gains over GPT-5.5 in OpenAI's startup spotlight series
MCP connector enables OAuth login with no API keys, from performance checks to cost estimates
Simon Willison review finds xhigh reasoning burns minutes even to draw a circle
Testing 45 agents on vulnerability hunting and game-building revealed a pattern where teamwork falls apart