Liquid AI unveils screen-reading model that runs in 3GB
LFM2.5-VL-3B hits 20 tokens/sec on Galaxy S26 Ultra, 228 tokens/sec on M5 Max
A company that makes the chips that power AI. Rather than a competitor in the chatbot race, it's the one selling picks and shovels to everyone in it. Its flagship products are GPUs and the AI infrastructure software built on top of them. In August 2026, it began officially supporting its cloud gaming service GeForce NOW on Linux, and around the same time, it published a method for isolating Kubernetes tenants on shared GPUs. In autonomous driving, it introduced a 34-billion-parameter reasoning model, and in robot policy learning, it presented a video-based learning approach. It also tried to expand the AI ecosystem for handling the physical world through its open world model 'Cosmos 3'.
Current rank (1M)
5
Rank over the last 30 days
2026.08.18 – 2026.09.16 · High 3 · Low 5 · Now 5
LFM2.5-VL-3B hits 20 tokens/sec on Galaxy S26 Ultra, 228 tokens/sec on M5 Max
IBM Cloud's first dedicated B300 cluster combined with Spectrum-X networking unveiled as inference infrastructure
The gap between companies that sell hardware and companies that profit from APIs over open-source strategy has become a talking point in the community
By running NVIDIA Nemotron 3.5 Lightning on local hardware, the company automates software development without exposing source code
Its accuracy was the lowest in its class, but its hallucination rate was also the lowest at 30%. The tradeoff: 50,000 tokens per answer.
An open-weight world model that produces a 10-second clip in 6.8 seconds, optimized for NVIDIA GPUs
Deal with six firms including Goldman Sachs, BlackRock to secure AI infrastructure funding; Bernstein and others warn of depreciation risk
The MoE architecture activates only 3B of its total 30B parameters per token, and NVIDIA claims up to 4x faster output than comparable models
Head of the General Intelligence Center answers directly on autonomous driving's limits and Europe entry timing
"DLR-Lock" research lets users run open-weight models as-is while blocking modification
DYNA-2, a robot foundation model pretrained on 170 years' worth of human behavior video, emerges
Muse Glimmer is an agent-specialized model that runs on 24GB-class consumer GPUs with 4-bit quantization