Lambda raises another $1 billion in private debt to buy NVIDIA chips
JPMorgan arranges short-term private notes so Lambda can secure GPUs to lease to Microsoft
An AI cloud that runs open-weight models on behalf of others, serving as the server room for companies wanting to use open-source models. Its flagship service is offering DeepSeek's latest model, V4 Pro, as an API, and it published its pricing sheet alongside the full rollout in August 2026. That same month, it expanded its infrastructure by supplying 10,000 NVIDIA B300 GPUs to India's largest AI data center. It was also revealed in August 2026 that, together with IBM and NVIDIA, it had built a B300 cluster on IBM Cloud. Around the same time, news emerged that Yutori's browser agent runs twice as fast as before on Together AI's environment.
Current rank (1M)
34
Rank over the last 30 days
2026.08.20 – 2026.09.18 · High 14 · Low 34 · Now 34
JPMorgan arranges short-term private notes so Lambda can secure GPUs to lease to Microsoft
Browser agent that repeats screen captures and clicks dozens of times cuts inference costs 4-5x
Partnership with Larsen & Toubro to build infrastructure for open-source inference, fine-tuning, and training
$0.435 input, $0.87 output per million tokens with 1M-token context — cache hits cut costs sharply
MAI Code 1.1 Flash improves on its predecessor but still trails the open-source DeepSeek V4 Flash
IBM Cloud's first dedicated B300 cluster combined with Spectrum-X networking unveiled as inference infrastructure
The gap between companies that sell hardware and companies that profit from APIs over open-source strategy has become a talking point in the community
The MoE architecture activates only 3B of its total 30B parameters per token, and NVIDIA claims up to 4x faster output than comparable models
Artificial Analysis to hold event in San Francisco on August 12 addressing speed gaps across inference providers
Muse Glimmer is an agent-specialized model that runs on 24GB-class consumer GPUs with 4-bit quantization
Moonshot AI tested Kimi K3 across major inference providers, with Together AI ranking first or tied for first in 3 of 4 benchmarks
Setup now allows assigning different open models to coding, planning, vision, and review stages