Home/Models

Model catalog

15 models from 7 providers, all reachable through the same /v1/chat/completions endpoint. Switch models with one parameter — no new contract, no new SDK.

15
Models published
7
Flagship tier
7
Upstream providers
1
API key to rotate
Flagship only

OpenAI · 4

GPT-4o 128K

Flagship

OpenAI's flagship multimodal model for text, vision and audio.

GPT-4o mini 128K

Cost-efficient small model for high-volume tasks.

o1 200K

Flagship

Reasoning model that thinks step by step on hard problems.

o3-mini 200K

Fast, affordable reasoning tuned for coding and STEM.

Anthropic · 3

Claude Opus 4 200K

Flagship

Anthropic's most capable model for complex agentic work.

Claude Sonnet 4 200K

Flagship

Balanced frontier intelligence and speed for production.

Claude Haiku 3.5 200K

Anthropic's fastest model for near-instant responses.

Google · 2

Gemini 2.5 Pro 1M

Flagship

Google's flagship model with a 1M-token context window.

Gemini 2.5 Flash 1M

Low-latency multimodal model at a fraction of the cost.

Meta · 2

Llama 4 Maverick 1M

Flagship

Meta's open-weight MoE model with 1M-token context.

Llama 3.3 70B 128K

The workhorse open-weight model for self-hosted workloads.

DeepSeek · 2

DeepSeek V3 64K

Strong open-weight generalist at commodity pricing.

DeepSeek R1 64K

Flagship

Open reasoning model competitive with frontier labs.

Mistral · 1

Mistral Large 2 128K

European flagship model with strong multilingual performance.

Qwen · 1

Qwen 2.5 72B 128K

Alibaba's multilingual model with excellent coding ability.

01 / FLAGSHIP TIER

The models most teams start with

Highest capability, longest context, and the ones our routing layer keeps the most warm capacity behind.

GPT-4o 128K

OpenAI

OpenAI's flagship multimodal model for text, vision and audio.

o1 200K

OpenAI

Reasoning model that thinks step by step on hard problems.

Claude Opus 4 200K

Anthropic

Anthropic's most capable model for complex agentic work.

Claude Sonnet 4 200K

Anthropic

Balanced frontier intelligence and speed for production.

Gemini 2.5 Pro 1M

Google

Google's flagship model with a 1M-token context window.

Llama 4 Maverick 1M

Meta

Meta's open-weight MoE model with 1M-token context.

DeepSeek R1 64K

DeepSeek

Open reasoning model competitive with frontier labs.

Need a model that is not listed? We onboard new vendors regularly and can route to private or self-hosted weights. Tell us what you need — most requests are answered within one business day.

Start routing in minutes

One endpoint. Every frontier model.

Create an account, pick a plan, and point your existing OpenAI client at our base URL. No SDK rewrite, no lock-in.