MiniMax H3 video generation now outpaces playback time
vLLM-Omni and FastVideo's FastH3 have pushed generation time below the video's own runtime
Hyunkook Kim
vLLM-Omni and FastVideo's FastH3 have pushed generation time below the video's own runtime
Same model, split by safeguard level — cheaper cache reads cut costs, and a new EFS system changes how data gets stored
On top of Anthropic's new-model price cut, Cognition's own harness delivers the same performance for 47% less
Tasks that start in the cloud automatically switch to Perplexity's own 27B model once they reach personal files
OpenAI has added hospital electronic health record integration and a plugin connecting nine official medical databases to ChatGPT for Healthcare.
The gap was 2.6x in January—now, half a year later, it's nearly tripled. What separates the two groups is how deeply they hand off real work to agents.
Top-tier hacking capability goes to defense partners first, with tighter restrictions for the general release
Google has added a new processing mode to Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite that lets the model pick out only the parts of a video it actually needs.
In a Fox Business interview, the NVIDIA CEO argued that data center construction is driving demand for skilled labor
AI solves diagnostic puzzles fast, but it can't judge the clues patients never say out loud, or the moment they're ready to hear the answer.
Cutting route selection from 20 minutes to 2.5 minutes, a last-mile carrier races ahead with NVIDIA AI
NVIDIA struck the deal with Lambda, a company it has backed, and took direct responsibility for leasing the Texas data center as well