METAL for iPhone

Read AI news in the METAL app.

Download METAL and discover fresh AI stories every day.

Download on the App Store

For iPhone · Free download

Search for METAL AI Magazine in the App Store on your iPhone.

METAL

AI GlossaryㅊInfrastructure and chips

Inference Provider

A company that supplies (or rents out) the infrastructure that actually runs the computations behind an AI model's responses.

In plain words

An Inference Provider is a company that handles the actual moment when an AI model computes an answer to a question. Think of it like a restaurant: the app in front of you is the storefront with the sign and the menu, taking orders from customers, while the Inference Provider is the kitchen behind the scenes that's rented out to actually cook the food. Customers order based on the storefront's name, but the kitchen that actually prepares the food is often run by a different company.

This setup means that even within a single app, different steps of a task can be handed off to different kitchens. For example, a coding tool might use different models for writing code, planning, looking at images, and reviewing results. Instead of relying on one all-purpose model, the app mixes and matches models that each have their own strengths for different steps.

The key point is that the company that built the app and the company that processes the inference can be different. Developers just need to choose and connect to whichever company's kitchen they want to use inside their app.

How it shows up in the news

Articles describe this as something like "supports Together AI as an inference provider," meaning Together AI supplies its own infrastructure to handle inference inside a coding agent tool like Roomote. A common misunderstanding is assuming the inference provider is the company that built the app. In reality, the team behind Roomote and Together AI, which handles inference, are separate companies.

See also

Stories using this term

Browse every entry