Twitter/X

Inference AutoRouter (announced by @samhogan on 2026-07-21) automatically…

Brief

Inference AutoRouter is a new routing layer that evaluates dozens of candidate models and directs each request to the best option based on cost/accuracy, exposing a single endpoint that can access hundreds of models. Developers enable it by appending ":auto" to a model name; the system entered private beta on 2026-07-21 (announcement by @samhogan).

Why it matters

Inference AutoRouter (announced by @samhogan on 2026-07-21) automatically evaluates dozens of models and routes requests based on cost vs. accuracy tradeoffs.

Key details

  • One single endpoint can access hundreds of models; enable routing by appending ":auto" to the model name.
  • The feature is available in private beta starting the announcement date.
Source evidence

We’re releasing Inference AutoRouter. Always the right model for your task

AutoRouter evals dozens of models and routes requests based on cost/accuracy

One endpoint. Hundreds of models

Append :auto to the model name to enable AutoRouter

Available in private beta today 👇