autoCapability-matched selection: a required-tier floor keeps only capable models, then a shrinkage-blended learned reward picks the cheapest of them. General workloads.
base_url.
Hand-curated catalog from Anthropic, OpenAI, Google, x-AI, DeepSeek,
Perplexity, Moonshot, Z.ai, and Groq — led by the Claude 5 family
(Opus 5, Sonnet 5, Fable 5). OpenAI- and Anthropic-compatible endpoints.
Set model="auto" and capability-matched
selection picks the best one per request.
Per-model quality, latency, and success variance are tracked live from real traffic. The leaderboard is public to every caller via GET /v1/intelligence/rankings — and GET /v1/catalog/runnable shows only what will actually route right now.
claude-opus-5 · claude-sonnet-5 · claude-fable-5 · claude-opus-4-7 · claude-haiku-4-5
gpt-5.6-sol · gpt-5.5 · gpt-4.1 · gpt-4o · o3 · o4-mini
gemini-3.5-flash · gemini-3.1-flash-lite · gemini-2.5-pro · gemini-2.5-flash
grok-4.6 · grok-4.5 · grok-4.3 · grok-4.20
deepseek-v4-pro · deepseek-v4-flash
sonar-pro · sonar-reasoning-pro · sonar-deep-research · sonar
kimi-k3 · kimi-k2.6 · kimi-k2.5
glm-5.2 · glm-5
openai/gpt-oss-120b
text-embedding-3-small · text-embedding-3-large
model to a strategy, not a name.autoCapability-matched selection: a required-tier floor keeps only capable models, then a shrinkage-blended learned reward picks the cheapest of them. General workloads.
auto:fastLowest p50 latency model. Real-time UX, streaming interfaces.
auto:floorCheapest model above the quality threshold. Bulk processing, classification, triage.
auto:bestHighest quality regardless of cost. Critical reasoning, code review, legal.
curl. Auto-routing across 40 models.