Chat & reasoning models — one API key, 35 models

Claude, GPT, Gemini, DeepSeek, Kimi, GLM, Qwen and Grok through the same OpenAI-compatible endpoint. Compare per-token prices side by side, then switch models by changing one string.

Crazyrouter is an AI API gateway that serves 35 chat and reasoning models through one OpenAI-compatible key, with per-token prices from $0.055 and native Anthropic, Gemini and Responses endpoints where the vendor has them.

35
models
$0.055
from, per 1M input
45%
max below list
Input price per 1M tokens (list vs Crazyrouter), cheapest first
ListCrazyrouterGemini 2.5 Flash-Lite$0.100 → $0.055GPT-4o mini$0.150 → $0.098Gemini 3 Flash$0.500 → $0.165Gemini 2.5 Flash$0.300 → $0.165DeepSeek V4 Flash$0.300 → $0.230GPT-4.1 mini$0.400 → $0.260MiniMax M3$0.300 → $0.300MiMo V2.6 Pro$0.435 → $0.435GPT-5.4 mini$0.750 → $0.488Kimi K2.5$0.600 → $0.600Claude Haiku 4.5$1.00 → $0.650DeepSeek V4 Pro$1.32 → $0.670

23 more models in the table below

What the models look like in use

All chat & reasoning models, by price

ModelVendorCrazyrouter price (per 1M tokens (in / out))ListContextDiscount
Gemini 2.5 Flash-Lite
gemini-2.5-flash-lite
Google$0.055 / $0.22 per 1M tokens$0.10 / $0.401M tokens−45%Try · Compare
GPT-4o mini
gpt-4o-mini
OpenAI$0.0975 / $0.39 per 1M tokens$0.15 / $0.60128K tokens−35%Try · Compare
Gemini 3 Flash
gemini-3-flash
Google$0.165 / $1.32 per 1M tokens$0.50 / $3.001M tokens−45%Try · Compare
Gemini 2.5 Flash
gemini-2.5-flash
Google$0.165 / $1.38 per 1M tokens$0.30 / $2.501M tokens−45%Try · Compare
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek$0.23 / $0.67 per 1M tokens$0.30 / $1.201M tokens—Try · Compare
GPT-4.1 mini
gpt-4.1-mini
OpenAI$0.26 / $1.04 per 1M tokens$0.40 / $1.601M tokens−35%Try · Compare
MiniMax M3
MiniMax-M3
MiniMax$0.30 / $1.20 per 1M tokens$0.30 / $1.201M tokens—Try · Compare
MiMo V2.6 Pro
mimo-v2.6-pro
Xiaomi$0.435 / $0.87 per 1M tokens$0.435 / $0.871M tokens—Try · Compare
GPT-5.4 mini
gpt-5.4-mini
OpenAI$0.488 / $2.93 per 1M tokens$0.75 / $4.50272K tokens−35%Try · Compare
Kimi K2.5
kimi-k2.5
Moonshot AI$0.60 / $3.15 per 1M tokens$0.60 / $3.00262K tokens—Try · Compare
Claude Haiku 4.5
claude-haiku-4-5
Anthropic$0.65 / $3.25 per 1M tokens$1.00 / $5.00200K tokens−35%Try · Compare
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek$0.67 / $2.00 per 1M tokens$1.32 / $3.961M tokens—Try · Compare
Gemini 2.5 Pro
gemini-2.5-pro
Google$0.688 / $5.50 per 1M tokens$1.25 / $10.001M tokens−45%Try · Compare
Gemini 3.8 Flash
gemini-3.8-flash
Google$0.75 / $3.75 per 1M tokens$0.75 / $3.751M tokens—Try · Compare
GPT-5
gpt-5
OpenAI$0.813 / $6.50 per 1M tokens$1.25 / $10.00272K tokens−35%Try · Compare
GPT-5 Codex
gpt-5-codex
OpenAI$0.813 / $6.50 per 1M tokens$1.25 / $10.00272K tokens−35%Try · Compare
GLM-5
glm-5
Zhipu AI$1.00 / $3.20 per 1M tokens$1.00 / $3.20200K tokens—Try · Compare
Gemini 3.1 Pro
gemini-3.1-pro
Google$1.10 / $6.60 per 1M tokens$2.00 / $12.001M tokens−45%Try · Compare
GPT-5.3 Codex
gpt-5.3-codex
OpenAI$1.14 / $9.10 per 1M tokens$1.75 / $14.00272K tokens−35%Try · Compare
Claude Sonnet 5
claude-sonnet-5
Anthropic$1.30 / $6.50 per 1M tokens$2.00 / $10.001M tokens−35%Try · Compare
GPT-4.1
gpt-4.1
OpenAI$1.30 / $5.20 per 1M tokens$2.00 / $8.001M tokens−35%Try · Compare
GPT-5.4
gpt-5.4
OpenAI$1.63 / $9.75 per 1M tokens$2.50 / $15.001M tokens−35%Try · Compare
GPT-4o
gpt-4o
OpenAI$1.63 / $6.50 per 1M tokens$2.50 / $10.00128K tokens−35%Try · Compare
Grok 4.6
grok-4.6
xAI$1.70 / $5.10 per 1M tokens$2.00 / $6.00500K tokens−15%Try · Compare
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic$1.95 / $9.75 per 1M tokens$3.00 / $15.00200K tokens−35%Try · Compare
OpenAI o3
o3
OpenAI$2.00 / $8.00 per 1M tokens$2.00 / $8.00200K tokens—Try · Compare
Qwen3.8 Max
qwen3.8-max
Alibaba Cloud$2.00 / $6.00 per 1M tokens$2.00 / $6.00992K tokens—Try · Compare
Kimi K3
kimi-k3
Moonshot AI$3.00 / $15.00 per 1M tokens$3.00 / $15.001M tokens—Try · Compare
Claude Opus 5
claude-opus-5
Anthropic$3.25 / $16.25 per 1M tokens$5.00 / $25.001M tokens−35%Try · Compare
Claude Opus 4.8
claude-opus-4-8
Anthropic$3.25 / $16.25 per 1M tokens$5.00 / $25.00200K tokens−35%Try · Compare
Claude Opus 4.7
claude-opus-4-7
Anthropic$3.25 / $16.25 per 1M tokens$5.00 / $25.00200K tokens−35%Try · Compare
GPT-5.5
gpt-5.5
OpenAI$3.25 / $19.50 per 1M tokens$5.00 / $30.001M tokens−35%Try · Compare
GPT-5.6 Sol
gpt-5.6-sol
OpenAI$3.25 / $19.50 per 1M tokens$4.00 / $20.00922K tokens−35%Try · Compare
GPT-6 Astra
gpt-6-astra
OpenAI$3.25 / $16.25 per 1M tokens$10.00 / $50.00922K tokens−35%Try · Compare
Claude Fable 5
claude-fable-5
Anthropic$6.50 / $32.50 per 1M tokens$10.00 / $50.001M tokens−35%Try · Compare

Prices come from the live billing table and already include the Crazyrouter discount; list prices are each vendor's public rate, for comparison only.

What is a chat & reasoning API?

An LLM API returns text (and increasingly tool calls and structured output) from a prompt. The models differ in reasoning depth, context window, speed and price; a gateway lets you route each request to the model that fits instead of integrating four vendors.

How to choose

Hard reasoning and coding

Claude Opus 5 / Opus 4.8, GPT-6 Astra, GPT-5.5, Gemini 3.1 Pro. Highest output price; reserve for tasks where quality changes the result.

Everyday assistants

Claude Sonnet 5 / 4.6, GPT-5.6 Sol, Gemini 3 Flash, DeepSeek V4 Flash, Kimi K3. Best price-to-quality for summaries, drafting, extraction.

High volume, low cost

Claude Haiku 4.5, GPT-5.4 mini, Gemini 2.5 Flash-Lite, GLM-5. Fractions of a cent per request; pair with caching.

Long context

Gemini 3.1 Pro and GPT-4.1 (1M tokens), Claude (200K). Check the context column before loading a whole repo.

One request shape for every model

Swap model for any row in the table; the example uses gemini-2.5-flash-lite.

cURL
curl https://api.crazyrouter.com/v1/chat/completions \
  -H "Authorization: Bearer $CRAZYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gemini-2.5-flash-lite","messages":[{"role":"user","content":"Summarize this PR in three bullets."}]}'

FAQ

What base URL do I use for chat models?

https://api.crazyrouter.com/v1 for OpenAI-compatible chat completions; Claude models also accept /v1/messages and Gemini models /v1beta.

Is the price the same as the vendor's?

No — the table shows the Crazyrouter price, which already includes the discount against each vendor's list price; the list price is shown for comparison.

Do streaming, tool calling and JSON mode work?

Yes, they behave as the upstream model documents them. Server-sent events are on by default when you pass stream: true.

Can I try a model before paying?

Every model page has a Try-it box, and new accounts get $2 of Playground credit; the API itself is pay as you go.

Which chat model is the cheapest on Crazyrouter?

Gemini 2.5 Flash-Lite at $0.055 per 1M input tokens as of the latest pricing sync; see the table for output prices.