Chat & reasoning models — one API key, 35 models
Claude, GPT, Gemini, DeepSeek, Kimi, GLM, Qwen and Grok through the same OpenAI-compatible endpoint. Compare per-token prices side by side, then switch models by changing one string.
Crazyrouter is an AI API gateway that serves 35 chat and reasoning models through one OpenAI-compatible key, with per-token prices from $0.055 and native Anthropic, Gemini and Responses endpoints where the vendor has them.
23 more models in the table below
What the models look like in use
All chat & reasoning models, by price
| Model | Vendor | Crazyrouter price (per 1M tokens (in / out)) | List | Context | Discount | |
|---|---|---|---|---|---|---|
| Gemini 2.5 Flash-Lite gemini-2.5-flash-lite | $0.055 / $0.22 per 1M tokens | $0.10 / $0.40 | 1M tokens | −45% | Try · Compare | |
| GPT-4o mini gpt-4o-mini | OpenAI | $0.0975 / $0.39 per 1M tokens | $0.15 / $0.60 | 128K tokens | −35% | Try · Compare |
| Gemini 3 Flash gemini-3-flash | $0.165 / $1.32 per 1M tokens | $0.50 / $3.00 | 1M tokens | −45% | Try · Compare | |
| Gemini 2.5 Flash gemini-2.5-flash | $0.165 / $1.38 per 1M tokens | $0.30 / $2.50 | 1M tokens | −45% | Try · Compare | |
| DeepSeek V4 Flash deepseek-v4-flash | DeepSeek | $0.23 / $0.67 per 1M tokens | $0.30 / $1.20 | 1M tokens | — | Try · Compare |
| GPT-4.1 mini gpt-4.1-mini | OpenAI | $0.26 / $1.04 per 1M tokens | $0.40 / $1.60 | 1M tokens | −35% | Try · Compare |
| MiniMax M3 MiniMax-M3 | MiniMax | $0.30 / $1.20 per 1M tokens | $0.30 / $1.20 | 1M tokens | — | Try · Compare |
| MiMo V2.6 Pro mimo-v2.6-pro | Xiaomi | $0.435 / $0.87 per 1M tokens | $0.435 / $0.87 | 1M tokens | — | Try · Compare |
| GPT-5.4 mini gpt-5.4-mini | OpenAI | $0.488 / $2.93 per 1M tokens | $0.75 / $4.50 | 272K tokens | −35% | Try · Compare |
| Kimi K2.5 kimi-k2.5 | Moonshot AI | $0.60 / $3.15 per 1M tokens | $0.60 / $3.00 | 262K tokens | — | Try · Compare |
| Claude Haiku 4.5 claude-haiku-4-5 | Anthropic | $0.65 / $3.25 per 1M tokens | $1.00 / $5.00 | 200K tokens | −35% | Try · Compare |
| DeepSeek V4 Pro deepseek-v4-pro | DeepSeek | $0.67 / $2.00 per 1M tokens | $1.32 / $3.96 | 1M tokens | — | Try · Compare |
| Gemini 2.5 Pro gemini-2.5-pro | $0.688 / $5.50 per 1M tokens | $1.25 / $10.00 | 1M tokens | −45% | Try · Compare | |
| Gemini 3.8 Flash gemini-3.8-flash | $0.75 / $3.75 per 1M tokens | $0.75 / $3.75 | 1M tokens | — | Try · Compare | |
| GPT-5 gpt-5 | OpenAI | $0.813 / $6.50 per 1M tokens | $1.25 / $10.00 | 272K tokens | −35% | Try · Compare |
| GPT-5 Codex gpt-5-codex | OpenAI | $0.813 / $6.50 per 1M tokens | $1.25 / $10.00 | 272K tokens | −35% | Try · Compare |
| GLM-5 glm-5 | Zhipu AI | $1.00 / $3.20 per 1M tokens | $1.00 / $3.20 | 200K tokens | — | Try · Compare |
| Gemini 3.1 Pro gemini-3.1-pro | $1.10 / $6.60 per 1M tokens | $2.00 / $12.00 | 1M tokens | −45% | Try · Compare | |
| GPT-5.3 Codex gpt-5.3-codex | OpenAI | $1.14 / $9.10 per 1M tokens | $1.75 / $14.00 | 272K tokens | −35% | Try · Compare |
| Claude Sonnet 5 claude-sonnet-5 | Anthropic | $1.30 / $6.50 per 1M tokens | $2.00 / $10.00 | 1M tokens | −35% | Try · Compare |
| GPT-4.1 gpt-4.1 | OpenAI | $1.30 / $5.20 per 1M tokens | $2.00 / $8.00 | 1M tokens | −35% | Try · Compare |
| GPT-5.4 gpt-5.4 | OpenAI | $1.63 / $9.75 per 1M tokens | $2.50 / $15.00 | 1M tokens | −35% | Try · Compare |
| GPT-4o gpt-4o | OpenAI | $1.63 / $6.50 per 1M tokens | $2.50 / $10.00 | 128K tokens | −35% | Try · Compare |
| Grok 4.6 grok-4.6 | xAI | $1.70 / $5.10 per 1M tokens | $2.00 / $6.00 | 500K tokens | −15% | Try · Compare |
| Claude Sonnet 4.6 claude-sonnet-4-6 | Anthropic | $1.95 / $9.75 per 1M tokens | $3.00 / $15.00 | 200K tokens | −35% | Try · Compare |
| OpenAI o3 o3 | OpenAI | $2.00 / $8.00 per 1M tokens | $2.00 / $8.00 | 200K tokens | — | Try · Compare |
| Qwen3.8 Max qwen3.8-max | Alibaba Cloud | $2.00 / $6.00 per 1M tokens | $2.00 / $6.00 | 992K tokens | — | Try · Compare |
| Kimi K3 kimi-k3 | Moonshot AI | $3.00 / $15.00 per 1M tokens | $3.00 / $15.00 | 1M tokens | — | Try · Compare |
| Claude Opus 5 claude-opus-5 | Anthropic | $3.25 / $16.25 per 1M tokens | $5.00 / $25.00 | 1M tokens | −35% | Try · Compare |
| Claude Opus 4.8 claude-opus-4-8 | Anthropic | $3.25 / $16.25 per 1M tokens | $5.00 / $25.00 | 200K tokens | −35% | Try · Compare |
| Claude Opus 4.7 claude-opus-4-7 | Anthropic | $3.25 / $16.25 per 1M tokens | $5.00 / $25.00 | 200K tokens | −35% | Try · Compare |
| GPT-5.5 gpt-5.5 | OpenAI | $3.25 / $19.50 per 1M tokens | $5.00 / $30.00 | 1M tokens | −35% | Try · Compare |
| GPT-5.6 Sol gpt-5.6-sol | OpenAI | $3.25 / $19.50 per 1M tokens | $4.00 / $20.00 | 922K tokens | −35% | Try · Compare |
| GPT-6 Astra gpt-6-astra | OpenAI | $3.25 / $16.25 per 1M tokens | $10.00 / $50.00 | 922K tokens | −35% | Try · Compare |
| Claude Fable 5 claude-fable-5 | Anthropic | $6.50 / $32.50 per 1M tokens | $10.00 / $50.00 | 1M tokens | −35% | Try · Compare |
Prices come from the live billing table and already include the Crazyrouter discount; list prices are each vendor's public rate, for comparison only.
What is a chat & reasoning API?
An LLM API returns text (and increasingly tool calls and structured output) from a prompt. The models differ in reasoning depth, context window, speed and price; a gateway lets you route each request to the model that fits instead of integrating four vendors.
How to choose
Claude Opus 5 / Opus 4.8, GPT-6 Astra, GPT-5.5, Gemini 3.1 Pro. Highest output price; reserve for tasks where quality changes the result.
Claude Sonnet 5 / 4.6, GPT-5.6 Sol, Gemini 3 Flash, DeepSeek V4 Flash, Kimi K3. Best price-to-quality for summaries, drafting, extraction.
Claude Haiku 4.5, GPT-5.4 mini, Gemini 2.5 Flash-Lite, GLM-5. Fractions of a cent per request; pair with caching.
Gemini 3.1 Pro and GPT-4.1 (1M tokens), Claude (200K). Check the context column before loading a whole repo.
One request shape for every model
Swap model for any row in the table; the example uses gemini-2.5-flash-lite.
curl https://api.crazyrouter.com/v1/chat/completions \
-H "Authorization: Bearer $CRAZYROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"gemini-2.5-flash-lite","messages":[{"role":"user","content":"Summarize this PR in three bullets."}]}'FAQ
What base URL do I use for chat models?
https://api.crazyrouter.com/v1 for OpenAI-compatible chat completions; Claude models also accept /v1/messages and Gemini models /v1beta.
Is the price the same as the vendor's?
No — the table shows the Crazyrouter price, which already includes the discount against each vendor's list price; the list price is shown for comparison.
Do streaming, tool calling and JSON mode work?
Yes, they behave as the upstream model documents them. Server-sent events are on by default when you pass stream: true.
Can I try a model before paying?
Every model page has a Try-it box, and new accounts get $2 of Playground credit; the API itself is pay as you go.
Which chat model is the cheapest on Crazyrouter?
Gemini 2.5 Flash-Lite at $0.055 per 1M input tokens as of the latest pricing sync; see the table for output prices.







