ant:bestQuality first — only models scoring 80 or above.
ant:best weights quality at the maximum and enforces a floor of 80, accepting higher latency and cost. Use it where the answer matters more than the bill.
Live routing over the last 30 days
Requests
181
Success rate
50.83%
Avg latency
565ms
73 models across 11 providers currently qualify. Per-request limits (context size, cost caps, your model allowlist) narrow this further at routing time.
o1-pro
o1-pro
Claude Sonnet 5
o3
o3
Anthropic: Claude Opus 4.7 (batch)
Anthropic: Claude Opus 4.1
Anthropic: Claude Opus 4.7
Anthropic: Claude Opus 4.8 (batch)
Anthropic: Claude Opus Latest
o1
Anthropic: Claude Opus 5.5 (batch)
Anthropic: Claude Opus 4.8
OpenAI: o4 Mini (batch)
Anthropic: Claude Opus 4.5 (batch)
Anthropic: Claude Opus 4.1 (batch)
Anthropic: Claude Opus 5 (batch)
Anthropic: Claude Opus 5
OpenAI: o4 Mini High
Anthropic: Claude Opus 4.6 (batch)
GPT-5
Anthropic: Claude Opus 4.5
Anthropic: Claude Opus 5.5
OpenAI: o3 Mini High
Showing the 24 highest-scoring of 73.
from openai import OpenAI
client = OpenAI(
base_url="https://api.antbase.ai/v1",
api_key="YOUR_ANT_API_KEY",
)
response = client.chat.completions.create(
model="ant:best",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Every virtual pool accepts the same request shape — swap the model id to change the routing policy.
Browse all virtual models →