Logical models

A logical model is the model name a client writes in its request, and it is where routing lands. Underneath it hangs a set of real provider models; once a request lands on it, one of them is chosen to actually run.

The Logical models page in the left sidebar manages request priority, switching mode, and model enablement.

One card, one logical model

Each logical model is a card:

  • Drag to reorder to set priority between logical models.
  • The model rows inside a card can be dragged too, deciding the try order during failover.
  • Add model creates a scheduling entry for the current logical model.

No explicit add, no participation

Only models explicitly added to a logical model take part in requests. Adding one creates a scheduling entry.

Two modes

Each logical model card has a toggle:

ModeBehavior
FailoverTries the attached models in order, automatically switching to the next on failure.
Manual pinningRequests use one fixed upstream model (that row is marked "pinned", the rest stand by).

Add a model before switching to manual

You cannot switch to manual pinning mode without any models.

Metric cards

The top of each card shows four metrics to help you judge who should rank first:

MetricMeaning
Success rateSuccess rate over the last N completed requests.
Average latencyFull latency of the attempt that served the request, including TTFT.
Average TPSOutput tokens ÷ full latency of that attempt (including TTFT).
Currently available modelsNumber of models available right now, and whether failover happened recently.

Each model row also shows: the time of its last success, its consecutive failure count, whether it is cooling down, and whether it is disabled.

The built-in default

default is the fallback logical model: any request that matches no other logical model lands here.

  • Its ID cannot be changed; its description still can.
  • It cannot be deleted.
  • It is the final safety net for every strategy in Request routing.

Create & edit

When you create a logical model, fill in:

  • Logical model ID: starts with a letter or digit and may contain letters, digits, dots, underscores, and hyphens, up to 64 characters. The model name in a request must be usable verbatim as this ID, so dots in version numbers are allowed (e.g. deepseek-v4.1-flash).
  • Description (optional).

The ID is the identity

Requests send it as the model name and the routing graph references it by it. Renaming swaps all of those references in the same save.

Delete a logical model

After deletion the UI no longer shows it, and requests that matched it fall to default.


Once logical models are arranged, go see how requests are actually routed.