AI models & prices

56 models

Every model below is reachable through one OpenAI-compatible endpoint and a single key. Pay per token — no subscriptions, no per-model accounts.

AnyModel puts the leading models from OpenAI, Anthropic, Google, DeepSeek and others behind a single OpenAI-compatible endpoint. Instead of juggling a separate account, key and bill for each provider, you reach all of them with one key and one balance.

Picking a model comes down to three things: capability, context window and price. Flagships like GPT-5.5, Claude Opus 4.8 and Gemini 3.1 Pro handle the hardest reasoning and agentic coding; lighter tiers such as Claude Haiku, GPT-5.4 mini and Gemini 3.1 Flash-Lite cost a fraction as much and suit high-volume or budget-sensitive work. Context windows run from 200K up to 1M tokens — the larger the window, the more code or documents a model can read in one pass.

The large figure on each card is what you actually pay per 1M tokens — final, with the model's multiplier already applied; metered routes show input and output rates separately. Next to it sits the vendor's own list price and how many times cheaper AnyModel is. You pay only for the tokens you actually use, with no subscription and no per-model minimums.

MiniMaxMiniMaxLive
MiniMax M3

Long-context multimodal reasoning at a budget price.

#38 · AI Arena
$0.00per 1M tokens512K ctx
base 5¢ ××0= $0.00
am/minimax-m3
OpenAIOpenAILive
GPT-5.4 mini

The strongest mini — cheap sub-agents and high-volume tasks.

$0.075per 1M tokens400K ctx
9.9× cheapervendor $0.75–4.5
base 5¢ ××1.5= $0.075
cx/gpt-5.4-mini
AnyModelAnyModelLive
Free Models Auto Router

Automatically selects a currently available free model.

$0.00per 1M tokensVaries ctx
base 5¢ ××0= $0.00
am/free
GoogleGoogleLive
Gemini 3.5 Flash

Google's newest default — frontier smarts at a Flash price.

#35 · AI Arena
$0.03per 1M tokens1M ctx
50× cheapervendor $1.5–9
base 5¢ ××0.6= $0.03
ag/gemini-3.5-flash-high
xAIxAIPopular
Grok 4.20 Multi-Agent Research

Grok's coordinated multi-agent mode for deep research.

$0.10per 1M tokens1M ctx
12× cheapervendor $1.25–2.5
base 5¢ ××2= $0.10
xai/grok-4.20-multi-agent-0309
UndisclosedLive
Ox Alpha

Free 1M-context reasoning model for long coding sessions.

$0.00per 1M tokens1M ctx
base 5¢ ××0= $0.00
am/ox-alpha
GoogleGoogleLive
Gemini 3.1 Flash-Lite

The fastest, cheapest Gemini of the 3rd series.

$0.03per 1M tokens1M ctx
8.3× cheapervendor $0.25–1.5
base 5¢ ××0.6= $0.03
ag/gemini-3.1-flash-lite-preview
NVIDIANVIDIALive
Nemotron 3 Ultra

NVIDIA's open reasoning flagship, available free through AnyModel.

#50 · AI Arena
$0.00per 1M tokens256K ctx
base 5¢ ××0= $0.00
am/nemotron-3-ultra-550b-a55b
GoogleGoogleLive
Gemma 4 31B

Google's open-weight model — near-free tokens for bulk work.

#52 · AI Arena
$0.00per 1M tokens128K ctx
base 5¢ ××0= $0.00
am/gemma-4-31b-it
GoogleGoogleLive
Gemini 2.5 Flash

Available through the live AnyModel catalog.

$0.03per 1M tokens1M ctx
base 5¢ ××0.6= $0.03
ag/gemini-2.5-flash
GoogleGoogleLive
Gemini 2.5 Flash Lite

Available through the live AnyModel catalog.

$0.03per 1M tokens1M ctx
base 5¢ ××0.6= $0.03
ag/gemini-2.5-flash-lite
GoogleGoogleLive
Gemini 2.5 Pro

Available through the live AnyModel catalog.

$0.085per 1M tokens1M ctx
base 5¢ ××1.7= $0.085
ag/gemini-2.5-pro
GoogleGoogleLive
Gemini 3 Flash

Available through the live AnyModel catalog.

#48 · AI Arena
$0.03per 1M tokens1M ctx
base 5¢ ××0.6= $0.03
ag/gemini-3-flash
GoogleGoogleLive
Gemini 3.6 Flash

Available through the live AnyModel catalog.

#39 · AI Arena
$0.03per 1M tokens1M ctx
base 5¢ ××0.6= $0.03
ag/gemini-3.6-flash-high
GoogleGoogleLive
Gemini 3.7 Flash

Available through the live AnyModel catalog.

#27 · AI Arena
$0.03per 1M tokens1M ctx
base 5¢ ××0.6= $0.03
ag/gemini-3.7-flash-high
GoogleGoogleLive
Gemini 3.1 Pro

Available through the live AnyModel catalog.

$0.085per 1M tokens1M ctx
base 5¢ ××1.7= $0.085
ag/gemini-pro-agent
OpenAIOpenAILive
GPT-OSS 120B

Available through the live AnyModel catalog.

$0.025per 1M tokens128K ctx
base 5¢ ××0.5= $0.025
ag/gpt-oss-120b-medium
GoogleGoogleLive
DiffusionGemma 26B A4B IT

Available through the live AnyModel catalog.

$0.00per 1M tokens128K ctx
base 5¢ ××0= $0.00
am/diffusiongemma-26b-a4b-it

The large figure is AnyModel's final price per 1M tokens — the model's multiplier is already in it; input/output rates are shown separately when they differ. The vendor's list price is shown next to it for comparison. USD.