Public access Live availability and current model status are shown per model. View status

Model access for production agents

The model you want.
The route that works.

Switch models, control cost, and keep one integration as supply changes.

Loading a recommended model…
One API keyUse the same integration across every supported model.
Health-aware routingInfer checks current route health and shows backup coverage when it exists.
Hard spend controlsSet limits before requests run and inspect every final charge.

Choose the outcome. Infer handles the route.

You decide what matters for the request. Infer shows the live availability and coverage behind each model.

01

Low cost

For background work, batch jobs, and flexible workloads.

03

Stable

Redundancy-oriented routing for work where continuity matters.

04

Official API

The model lab’s first-party endpoint, accessed through Infer at a fixed 1.25× supplier rate.

Why not use every model API directly?

Direct APIs are excellent when one model and one provider are enough. Infer is for teams that want to compare models, enforce spend limits, and change supply without rebuilding the integration.

With separate APIsWith Infer
Different keys and request formatsOne API and one key
Prices checked across separate dashboardsCurrent Infer prices in one catalog
Spend controls built separatelyLimits enforced on every key
Availability changes across separate APIsCurrent model status in one catalog

One integration. Clear availability.

Every model shows current availability, capabilities, and current pricing so you can choose without relying on hidden supply details.

How routing works
Capabilities shownSee current context, tools, structured-output, streaming, and availability information.
Receipt on every callTokens, selected option, and final charge stay visible.

Pick a model. Send the request.

Every option uses the same API and the same controls.