Liquid AI's 300M Draft Model Speeds Up Decoding by Up to 3.18x
DSpark released for three LFM2.5 models — greedy output identical to the original, MacBook throughput up 61→139 tokens/sec
A public marketplace where AI models are gathered and distributed. Rather than a company that builds models itself, it functions as "the GitHub of models," hosting publicly available models built by others. The Hub hosts a wide range of open-weight models, from transformer-based models to image and speech generation models. The structure is such that various providers plug in inference infrastructure for users who want to download and run models themselves, and in August 2026, Baseten joined this list as an inference provider for Hugging Face.
Current rank (1M)
6
Rank over the last 30 days
2026.08.17 – 2026.09.15 · High 3 · Low 6 · Now 6
DSpark released for three LFM2.5 models — greedy output identical to the original, MacBook throughput up 61→139 tokens/sec
It won even in a test that standardizes voices and pits only the engines against each other, using an SSM architecture instead of a transformer
In IBM Research's tests across 8 models, DeepSeek preferred full guideline sets while gpt-oss performed better with compressed versions
Liquid AI added a loop, and both coding agents completed the production-grade task
New monitoring aims to flag alerts within 30 minutes, but largest frontier RL training remains paused
Six months with the $899 AI robot litter box — waste removal was excellent, but facial recognition and waste detection fell short
Open-weight MiniMax-Music3 generates a complete 32kHz stereo track in one pass from just two inputs: lyrics and a structured caption
Former OpenAI staffer Brundage tells Bloomberg third-party audits are needed
LabLLM supports GPT-style model design, training, and chat locally on Apple Silicon via MLX
New custom benchmarking platform compares not just quality but cost and time per task
The open-weight model, released on the 14th, topped the charts in downloads and likes
Hugging Face's "State of Open Models" report shows Qwen leading Gemma in real-world local inference use