METAL

Lovable and Cerebras Partner on Infrastructure to Speed Up AI Response Times

Lovable is teaming up with Cerebras to build infrastructure that will dramatically cut response latency by 2027. Since launching in November 2024, the platform has surpassed 50 million cumulative projects.

Lovable and Cerebras Partner on Infrastructure to Speed Up AI Response Times

Summary

  • Lovable has partnered with Cerebras to gradually roll out dedicated infrastructure that speeds up AI response times.
  • Cerebras's wafer-scale engine places an entire model on a single silicon wafer, eliminating inter-chip communication overhead.
  • The rollout will begin with latency-sensitive workloads first, with progress to be disclosed along the way.

50 Million Projects and the Wall of "Latency"

Lovable is a platform for building apps and websites through conversation with AI. Users describe what they want, and the AI generates a working prototype in real time, which can then be refined through simple feedback and deployed. Since launching in November 2024, the number of projects built on Lovable has surpassed 50 million.

The platform's core loop is "describe, see the result, iterate." But when waiting time creeps in between each step, the workflow breaks down. This is precisely why Lovable decided to invest in infrastructure—to cut down on that "waiting time."

How Cerebras's Wafer-Scale Engine Works

Typical AI systems spread a model's memory across dozens of separate chips. Every time those chips exchange information with one another, it incurs a "communication cost," which is one source of response latency. Cerebras took a different approach. Its Wafer-Scale Engine places an entire model on a single, massive silicon wafer, eliminating inter-chip communication altogether.

Cerebras says this architecture is especially effective in environments with repeated back-and-forth exchanges, like software development, and that it accelerates the slowest part of the process alongside existing GPUs. Andrew Feldman, co-founder and CEO of Cerebras, said, "When AI responds in real time, users do more, stay longer, and complete more valuable work."

Phased Rollout and Target Timeline

Lovable is aiming to dramatically cut response times by 2027. Rather than an immediate, full-scale switch, the company has chosen a phased approach—migrating the most latency-sensitive workloads to dedicated Cerebras capacity first. Progress on each phase of the transition will be disclosed sequentially.

Anton Osika, co-founder and CEO of Lovable, said, "We built Lovable on the idea that the person closest to a problem should be able to solve it themselves," adding, "Cerebras helps Lovable respond as fast as our customers can think."

Pricing, Scope of Support, and Limitations

The current information does not include specific details on pricing changes, billing structure, or the scope of supported regions and models related to this infrastructure transition. Nor has a timeline been disclosed for when the full migration to dedicated Cerebras capacity will be complete, or for each phase's schedule. The company has stated that progress updates will be shared at a later date.

Comments