Estimate OpenRouter costs from the selected route, supported charge categories and routing constraints. Separate model usage from account-credit operations and any BYOK arrangement.

How OpenRouter pricing works

The model schema records token, request, media and cache categories, with conditional pricing overrides when an endpoint uses them. Official documentation.

Begin with the complete operation your application performs. A plain text completion, a search-enabled response and a media-generation task may involve different billable work. Preserve the unit beside every amount. Converting all rows into an undifferentiated token figure would hide the condition that makes the estimate meaningful.

The default catalog price is a starting reference, not a guarantee that every possible endpoint configuration produces the same amount. Keep routing choices in the scenario and inspect the eligible endpoint evidence when a narrow price difference determines the decision. A maximum-price condition can reduce eligibility as well as constrain cost.

BYOK connects your provider credentials through OpenRouter and has its own documented fee and fallback conditions. Official documentation.

Keep direct-provider spending and OpenRouter-account spending visible as separate records when using that mode. A familiar model charge appearing in two dashboards needs reconciliation, not an assumption that one is a duplicate. Document which account pays for inference and which settings permit a fallback to another billing route.

PNG showing selected model/endpoint, ordinary and cached usage, optional units, BYOK/account split and routing price constraint.

Full price list

Verified model prices
ModelPrice typeUSDUnitTierRegionSource
AionLabs: Aion-2.0Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-2.0Input$0.8per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-2.0Output$1.6per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0Cached Input$0.75per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0Input$3per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0Output$6per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0-MiniCached Input$0.18per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0-MiniInput$0.7per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-3.0-MiniOutput$1.4per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-RP 1.0 (8B)Input$0.8per 1M tokensStandardSee sourceOfficial source ↗
AionLabs: Aion-RP 1.0 (8B)Output$1.6per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova 2 LiteInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova 2 LiteOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Lite 1.0Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Lite 1.0Output$0.24per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Micro 1.0Input$0.035per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Micro 1.0Output$0.14per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Premier 1.0Cached Input$0.625per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Premier 1.0Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Premier 1.0Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Pro 1.0Input$0.8per 1M tokensStandardSee sourceOfficial source ↗
Amazon: Nova Pro 1.0Output$3.2per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude 3 HaikuCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude 3 HaikuInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude 3 HaikuOutput$1.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5Cached Input$1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5Input$10per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5Output$50per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5 (batch)Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5 (batch)Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5 (batch)Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1Input$10per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1Output$50per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1 (batch)Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1 (batch)Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable 5.1 (batch)Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable LatestCached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable LatestInput$10per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Fable LatestOutput$50per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5Input$1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5Output$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5 (batch)Cached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5 (batch)Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku 4.5 (batch)Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku LatestCached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku LatestInput$1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Haiku LatestOutput$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4Cached Input$1.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4Input$15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4Output$75per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1Cached Input$1.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1Input$15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1Output$75per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1 (batch)Cached Input$0.75per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1 (batch)Input$7.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.1 (batch)Output$37.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.5 (batch)Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.6 (batch)Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.7 (batch)Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8Input$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8Output$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus 4.8 (batch)Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus LatestCached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus LatestInput$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Opus LatestOutput$25per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4Cached Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4Input$3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4Output$15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5Cached Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5Input$3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5Output$15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5 (batch)Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5 (batch)Input$1.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.5 (batch)Output$7.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6Cached Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6Input$3per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6Output$15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6 (batch)Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6 (batch)Input$1.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 4.6 (batch)Output$7.5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5Input$2per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5Output$10per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5 (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5 (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet 5 (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet LatestCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet LatestInput$2per 1M tokensStandardSee sourceOfficial source ↗
Anthropic: Claude Sonnet LatestOutput$10per 1M tokensStandardSee sourceOfficial source ↗
Arcee AI: Trinity Large ThinkingCached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
Arcee AI: Trinity Large ThinkingInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Arcee AI: Trinity Large ThinkingOutput$0.8per 1M tokensStandardSee sourceOfficial source ↗
Baidu: ERNIE 4.5 VL 424B A47B Input$0.42per 1M tokensStandardSee sourceOfficial source ↗
Baidu: ERNIE 4.5 VL 424B A47B Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 1.6Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 1.6Output$2per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 1.6 FlashInput$0.075per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 1.6 FlashOutput$0.3per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 2.1 TurboInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed 2.1 TurboOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-CodeInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-CodeOutput$3per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-LiteInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-LiteOutput$2per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-MiniInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
ByteDance Seed: Seed-2.0-MiniOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
ByteDance: UI-TARS 7B Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
ByteDance: UI-TARS 7B Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
ByteDance: UI-TARS 7B Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5Input$5per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5Output$25per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Claude Opus 5 (batch)Output$12.5per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command AInput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command AOutput$10per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R (08-2024)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R (08-2024)Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R+ (08-2024)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R+ (08-2024)Output$10per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R7B (12-2024)Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
Cohere: Command R7B (12-2024)Output$0.15per 1M tokensStandardSee sourceOfficial source ↗
Cohere: North Mini Code (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Cohere: North Mini Code (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3Input$0.2574per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3Output$1.03per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3 0324Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3 0324Output$1per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1Cached Input$0.13per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1Output$0.95per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1 TerminusCached Input$0.135per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1 TerminusInput$0.27per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.1 TerminusOutput$1per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.2Cached Input$0.1345per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.2Input$0.269per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.2Output$0.4per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.2 ExpInput$0.27per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V3.2 ExpOutput$0.41per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0423Cached Input$0.0132per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0423Input$0.0659per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0423Output$0.1319per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731Cached Input$0.008per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731Output$0.08per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731 (batch)Cached Input$0.0035per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731 (batch)Input$0.11per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash 0731 (batch)Output$0.33per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash LatestCached Input$0.003per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash LatestInput$0.03per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash LatestOutput$0.07per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision ExpCached Input$0.007per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision ExpInput$0.22per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision ExpOutput$0.66per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision Exp (batch)Cached Input$0.0035per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision Exp (batch)Input$0.11per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Flash Vision Exp (batch)Output$0.33per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0423Cached Input$0.0615per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0423Input$0.7383per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0423Output$1.48per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813Cached Input$0.0193per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813Input$0.5795per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813Output$1.74per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813 (batch)Cached Input$0.022per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813 (batch)Input$0.66per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4 Pro 0813 (batch)Output$1.98per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4.1 FlashCached Input$0.003per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4.1 FlashInput$0.15per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: DeepSeek V4.1 FlashOutput$0.6per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1Input$0.7per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1 0528Cached Input$0.35per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1 0528Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1 0528Output$2.15per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1 Distill Llama 70BInput$0.8per 1M tokensStandardSee sourceOfficial source ↗
DeepSeek: R1 Distill Llama 70BOutput$0.8per 1M tokensStandardSee sourceOfficial source ↗
Dots Studio: Dots3-Note Preview (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Dots Studio: Dots3-Note Preview (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Free Models RouterInputFreeper 1M tokensStandardSee sourceOfficial source ↗
Free Models RouterOutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 FlashCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 FlashInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 FlashOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 FlashPer Image$0.0000003per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash (batch)Cached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash (batch)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash (batch)Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash (batch)Per Image$0.00000015per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash LiteCached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash LiteInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash LiteOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash LitePer Image$0.0000001per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash Lite (batch)Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash Lite (batch)Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash Lite (batch)Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Flash Lite (batch)Per Image$0.00000005per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 ProCached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 ProInput$1.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 ProOutput$10per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 ProPer Image$0.000001per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro (batch)Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro (batch)Input$0.625per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro (batch)Per Image$0.000000625per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 05-06Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 05-06Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 05-06Output$10per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 05-06Per Image$0.000001per imageStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 06-05Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 06-05Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 06-05Output$10per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 2.5 Pro Preview 06-05Per Image$0.000001per imageStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash PreviewCached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash PreviewInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash PreviewOutput$3per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash PreviewPer Image$0.0000005per imageStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash Preview (batch)Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash Preview (batch)Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3 Flash Preview (batch)Per Image$0.00000025per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash LiteCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash LiteInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash LiteOutput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash LitePer Image$0.00000025per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite (batch)Cached Input$0.0125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite (batch)Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite (batch)Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite (batch)Per Image$0.000000125per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite PreviewCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite PreviewInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite PreviewOutput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Flash Lite PreviewPer Image$0.00000025per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro PreviewCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro PreviewInput$2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro PreviewOutput$12per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro PreviewPer Image$0.000002per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview (batch)Output$6per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview (batch)Per Image$0.000001per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview Custom ToolsCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview Custom ToolsInput$2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview Custom ToolsOutput$12per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.1 Pro Preview Custom ToolsPer Image$0.000002per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.5 FlashCached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 FlashInput$1.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 FlashOutput$9per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 FlashPer Image$0.000002per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash (batch)Cached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash (batch)Input$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash (batch)Output$4.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash (batch)Per Image$0.00000075per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash LiteCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash LiteInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash LiteOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash LitePer Image$0.0000003per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash Lite (batch)Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash Lite (batch)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash Lite (batch)Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.5 Flash Lite (batch)Per Image$0.00000015per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.6 FlashCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 FlashInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 FlashOutput$3.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 FlashPer Image$0.00000075per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.6 Flash (batch)Cached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 Flash (batch)Input$0.375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 Flash (batch)Output$1.88per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.6 Flash (batch)Per Image$0.000000375per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.7 FlashCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 FlashInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 FlashOutput$3.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 FlashPer Image$0.00000075per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.7 Flash (batch)Cached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 Flash (batch)Input$0.375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 Flash (batch)Output$1.88per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.7 Flash (batch)Per Image$0.000000375per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.8 FlashCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 FlashInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 FlashOutput$3.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 FlashPer Image$0.00000075per imageStandardSee sourceOfficial source ↗
Google: Gemini 3.8 Flash (batch)Cached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 Flash (batch)Input$0.375per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 Flash (batch)Output$1.88per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini 3.8 Flash (batch)Per Image$0.000000375per imageStandardSee sourceOfficial source ↗
Google: Gemini Flash LatestCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Flash LatestInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Flash LatestOutput$3.75per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Flash LatestPer Image$0.00000075per imageStandardSee sourceOfficial source ↗
Google: Gemini Pro LatestCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Pro LatestInput$2per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Pro LatestOutput$12per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemini Pro LatestPer Image$0.000002per imageStandardSee sourceOfficial source ↗
Google: Gemma 2 27BInput$0.65per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 2 27BOutput$0.65per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 12BInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 12BOutput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 27BCached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 27BInput$0.08per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 27BOutput$0.45per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 4BInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 3 4BOutput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 26B A4B Input$0.042per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 26B A4B Output$0.22per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 26B A4B (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 26B A4B (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31BCached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31BInput$0.09per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31BOutput$0.34per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31B (batch)Input$0.39per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31B (batch)Output$0.97per 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31B (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Gemma 4 31B (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Lyria 3 Clip PreviewInputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Lyria 3 Clip PreviewOutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Lyria 3 Pro PreviewInputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Lyria 3 Pro PreviewOutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana (Gemini 2.5 Flash Image)Cached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana (Gemini 2.5 Flash Image)Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana (Gemini 2.5 Flash Image)Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana (Gemini 2.5 Flash Image)Per Image$0.0000003per imageStandardSee sourceOfficial source ↗
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)Output$3per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana 2 (Gemini 3.1 Flash Image)Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana 2 (Gemini 3.1 Flash Image)Output$3per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)Input$2per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)Output$12per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)Per Image$0.000002per imageStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image)Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image)Input$2per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image)Output$12per 1M tokensStandardSee sourceOfficial source ↗
Google: Nano Banana Pro (Gemini 3 Pro Image)Per Image$0.000002per imageStandardSee sourceOfficial source ↗
IBM: Granite 4.0 MicroInput$0.017per 1M tokensStandardSee sourceOfficial source ↗
IBM: Granite 4.0 MicroOutput$0.112per 1M tokensStandardSee sourceOfficial source ↗
IBM: Granite 4.2 8BCached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
IBM: Granite 4.2 8BInput$0.06per 1M tokensStandardSee sourceOfficial source ↗
IBM: Granite 4.2 8BOutput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2Cached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2.5Cached Input$0.004per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2.5Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Inception: Mercury 2.5Output$0.15per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 FlashCached Input$0.0042per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 FlashInput$0.021per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 FlashOutput$0.063per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash FinCached Input$0.012per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash FinInput$0.06per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash FinOutput$0.18per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash Fin (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash Fin (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash Sante (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash Sante (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash VLCached Input$0.012per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash VLInput$0.06per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash VLOutput$0.18per 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash VL (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
inclusionAI: Ling 3.0 Flash VL (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Inference.net: Schematron V2 SmallInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Inference.net: Schematron V2 SmallOutput$0.23per 1M tokensStandardSee sourceOfficial source ↗
Inference.net: Schematron V2 TurboInput$0.03per 1M tokensStandardSee sourceOfficial source ↗
Inference.net: Schematron V2 TurboOutput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2Cached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2.5Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2.5Input$0.74per 1M tokensStandardSee sourceOfficial source ↗
Kwaipilot: KAT-Coder-Pro V2.5Output$2.96per 1M tokensStandardSee sourceOfficial source ↗
LiquidAI: LFM2.5-2.6B (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
LiquidAI: LFM2.5-2.6B (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Magnum v4 72BInput$2.5per 1M tokensStandardSee sourceOfficial source ↗
Magnum v4 72BOutput$5per 1M tokensStandardSee sourceOfficial source ↗
Mancer: Weaver (alpha)Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Mancer: Weaver (alpha)Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Meituan: LongCat 2.0Cached Input$0.006per 1M tokensStandardSee sourceOfficial source ↗
Meituan: LongCat 2.0Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Meituan: LongCat 2.0Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.1 70B InstructInput$0.72per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.1 70B InstructOutput$0.72per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.1 8B InstructCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.1 8B InstructInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.1 8B InstructOutput$0.08per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.2 1B InstructInput$0.027per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.2 1B InstructOutput$0.201per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.2 3B InstructInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.2 3B InstructOutput$0.33per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.3 70B InstructInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 3.3 70B InstructOutput$0.32per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 4 MaverickInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 4 MaverickOutput$0.696per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 4 ScoutInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama 4 ScoutOutput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama Guard 4 12BInput$0.18per 1M tokensStandardSee sourceOfficial source ↗
Meta: Llama Guard 4 12BOutput$0.18per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30BCached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30BInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30BOutput$1.1per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30B (batch)Cached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30B (batch)Input$0.175per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Glimmer 30B (batch)Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.1Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.1Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.1Output$4.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2Output$4.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2 ContributorCached Input$0.002per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2 ContributorInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.2 ContributorOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3Output$4.25per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3 ContributorCached Input$0.002per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3 ContributorInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Meta: Muse Spark 1.3 ContributorOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Microsoft: Phi 4Input$0.07per 1M tokensStandardSee sourceOfficial source ↗
Microsoft: Phi 4Output$0.14per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M1Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M1Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2Input$0.255per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2Output$1.02per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2-herCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2-herInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2-herOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.1Cached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.1Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.1Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.5Cached Input$0.027per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.5Input$0.27per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.5Output$1.08per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.7Cached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.7Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M2.7Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3Cached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3 (batch)Cached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3 (batch)Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax M3 (batch)Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax-01Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
MiniMax: MiniMax-01Output$1.1per 1M tokensStandardSee sourceOfficial source ↗
Mistral LargeCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral LargeInput$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral LargeOutput$6per 1M tokensStandardSee sourceOfficial source ↗
Mistral Large 2407Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral Large 2407Input$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral Large 2407Output$6per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508Cached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508Output$0.9per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508 (batch)Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508 (batch)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Codestral 2508 (batch)Output$0.45per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Devstral 2 2512Cached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Devstral 2 2512Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Devstral 2 2512Output$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 14B 2512Cached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 14B 2512Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 14B 2512Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 3B 2512Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 3B 2512Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 3B 2512Output$0.1per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512Output$0.15per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512 (batch)Cached Input$0.0075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512 (batch)Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Ministral 3 8B 2512 (batch)Output$0.075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512Cached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512 (batch)Cached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512 (batch)Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Large 3 2512 (batch)Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3Cached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3Output$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1Cached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1Output$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1 (batch)Cached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1 (batch)Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.1 (batch)Output$1per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.5Input$1.5per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.5Output$7.5per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.5 (batch)Input$0.75per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Medium 3.5 (batch)Output$3.75per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral NemoInput$0.019per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral NemoOutput$0.03per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3Output$0.08per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3.1 24BInput$0.351per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3.1 24BOutput$0.555per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3.2 24BInput$0.075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 3.2 24BOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4 (batch)Cached Input$0.0075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4 (batch)Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mistral Small 4 (batch)Output$0.3per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mixtral 8x22B InstructCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mixtral 8x22B InstructInput$2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Mixtral 8x22B InstructOutput$6per 1M tokensStandardSee sourceOfficial source ↗
Mistral: SabaCached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
Mistral: SabaInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Mistral: SabaOutput$0.6per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Voxtral Small 24B 2507Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Voxtral Small 24B 2507Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Mistral: Voxtral Small 24B 2507Output$0.3per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 0711Input$0.57per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 0711Output$2.3per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 0905Input$0.6per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 0905Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 ThinkingInput$0.6per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2 ThinkingOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.5Cached Input$0.07per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.5Input$0.45per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.5Output$2.25per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.6Cached Input$0.16per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.6Input$0.95per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.6Output$4per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.7 CodeCached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.7 CodeInput$0.71per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K2.7 CodeOutput$3.5per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3Cached Input$0.2632per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3Input$2.3per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3Output$11.55per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3 (batch)Cached Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3 (batch)Input$3per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi K3 (batch)Output$15per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi LatestCached Input$0.2465per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi LatestInput$2.13per 1M tokensStandardSee sourceOfficial source ↗
MoonshotAI: Kimi LatestOutput$11.9per 1M tokensStandardSee sourceOfficial source ↗
Morph: Morph V3 FastInput$0.8per 1M tokensStandardSee sourceOfficial source ↗
Morph: Morph V3 FastOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
Morph: Morph V3 LargeInput$0.9per 1M tokensStandardSee sourceOfficial source ↗
Morph: Morph V3 LargeOutput$1.9per 1M tokensStandardSee sourceOfficial source ↗
MythoMax 13BInput$0.06per 1M tokensStandardSee sourceOfficial source ↗
MythoMax 13BOutput$0.06per 1M tokensStandardSee sourceOfficial source ↗
Nex AGI: Nex-N2.5-Mini (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Nex AGI: Nex-N2.5-Mini (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Nex AGI: Nex-N2.5-Pro (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Nex AGI: Nex-N2.5-Pro (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 3 405B InstructInput$1per 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 3 405B InstructOutput$1per 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 3 70B InstructInput$0.7per 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 3 70B InstructOutput$0.7per 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 4 405BInput$1per 1M tokensStandardSee sourceOfficial source ↗
Nous: Hermes 4 405BOutput$3per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Nano 30B A3BCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Nano 30B A3BInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Nano 30B A3BOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Nano Omni (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Nano Omni (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 SuperInput$0.085per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 SuperOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Super (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Super (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 UltraCached Input$0.1875per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 UltraInput$0.625per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 UltraOutput$3.13per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Ultra (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3 Ultra (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Content SafetyInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Content SafetyOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Content Safety (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Content Safety (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 LightningCached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 LightningInput$0.08per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 LightningOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Lightning (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
NVIDIA: Nemotron 3.5 Lightning (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Astra LatestCached Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Astra LatestInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Astra LatestOutput$50per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT AudioInput$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT AudioOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Audio MiniInput$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Audio MiniOutput$2.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Chat LatestCached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Chat LatestInput$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Chat LatestOutput$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Luna LatestCached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Luna LatestInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Luna LatestOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Mini LatestCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Mini LatestInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Mini LatestOutput$4.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Sol LatestCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Sol LatestInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Sol LatestOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Terra LatestCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Terra LatestInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT Terra LatestOutput$12per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 TurboInput$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 TurboOutput$1.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo (batch)Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo (batch)Output$0.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo (older v0613)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo (older v0613)Output$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo 16kInput$3per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo 16kOutput$4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo InstructInput$1.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-3.5 Turbo InstructOutput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4Input$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4Output$60per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 TurboInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 TurboOutput$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 Turbo (batch)Input$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 Turbo (batch)Output$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 Turbo PreviewInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4 Turbo PreviewOutput$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1Input$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1Output$8per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 (batch)Output$4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 MiniCached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 MiniInput$0.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 MiniOutput$1.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Mini (batch)Cached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Mini (batch)Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Mini (batch)Output$0.8per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 NanoCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 NanoInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 NanoOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Nano (batch)Cached Input$0.0125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Nano (batch)Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4.1 Nano (batch)Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4oCached Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4oInput$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4oOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-05-13)Input$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-05-13)Output$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-08-06)Cached Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-08-06)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-08-06)Output$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-11-20)Cached Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-11-20)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (2024-11-20)Output$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (batch)Cached Input$0.625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (batch)Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-miniCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-miniInput$0.15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-miniOutput$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (2024-07-18)Cached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (2024-07-18)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (2024-07-18)Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (batch)Cached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (batch)Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-4o-mini (batch)Output$0.3per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5Output$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 (batch)Cached Input$0.0625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 (batch)Input$0.625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 ImageCached Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 ImageInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 ImageOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Image MiniCached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Image MiniInput$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Image MiniOutput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 MiniCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 MiniInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 MiniOutput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Mini (batch)Cached Input$0.0125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Mini (batch)Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Mini (batch)Output$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 NanoCached Input$0.005per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 NanoInput$0.05per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 NanoOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Nano (batch)Cached Input$0.0025per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Nano (batch)Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Nano (batch)Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 ProInput$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 ProOutput$120per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Pro (batch)Input$7.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5 Pro (batch)Output$60per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1Output$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1 (batch)Cached Input$0.0625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1 (batch)Input$0.625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1 (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-CodexCached Input$0.13per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-CodexInput$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-CodexOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MaxCached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MaxInput$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MaxOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MiniCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MiniInput$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.1-Codex-MiniOutput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2Cached Input$0.175per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2Input$1.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2Output$14per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 (batch)Cached Input$0.0875per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 (batch)Input$0.875per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 (batch)Output$7per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 ChatCached Input$0.175per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 ChatInput$1.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 ChatOutput$14per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 ProInput$21per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 ProOutput$168per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 Pro (batch)Input$10.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2 Pro (batch)Output$84per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2-CodexCached Input$0.175per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2-CodexInput$1.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.2-CodexOutput$14per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.3-CodexCached Input$0.175per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.3-CodexInput$1.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.3-CodexOutput$14per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4Output$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 (batch)Cached Input$0.125per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 (batch)Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 (batch)Output$7.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Image 2Cached Input$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Image 2Input$8per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Image 2Output$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 MiniCached Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 MiniInput$0.75per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 MiniOutput$4.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Mini (batch)Cached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Mini (batch)Input$0.375per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Mini (batch)Output$2.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 NanoCached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 NanoInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 NanoOutput$1.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Nano (batch)Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Nano (batch)Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Nano (batch)Output$0.625per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 ProInput$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 ProOutput$180per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Pro (batch)Input$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.4 Pro (batch)Output$90per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5Input$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5Output$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 (batch)Input$2.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 (batch)Output$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 ProInput$30per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 ProOutput$180per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 Pro (batch)Input$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.5 Pro (batch)Output$90per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 LunaCached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 LunaInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 LunaOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna (batch)Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna (batch)Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna (batch)Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna ProCached Input$0.02per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna ProInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna ProOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna Pro (batch)Cached Input$0.01per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna Pro (batch)Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Luna Pro (batch)Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 SolCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 SolInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 SolOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol ProCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol ProInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol ProOutput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol Pro (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol Pro (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Sol Pro (batch)Output$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 TerraCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 TerraInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 TerraOutput$12per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra (batch)Output$6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra ProCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra ProInput$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra ProOutput$12per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra Pro (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra Pro (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-5.6 Terra Pro (batch)Output$6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 AstraCached Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 AstraInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 AstraOutput$50per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra (batch)Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra (batch)Input$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra (batch)Output$25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra ProCached Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra ProInput$10per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra ProOutput$50per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra Pro (batch)Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra Pro (batch)Input$5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: GPT-6 Astra Pro (batch)Output$25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-120bInput$0.037per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-120bOutput$0.17per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-120b (batch)Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-120b (batch)Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-20bCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-20bInput$0.03per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-20bOutput$0.13per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-20b (batch)Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-20b (batch)Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-safeguard-20bCached Input$0.0375per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-safeguard-20bInput$0.075per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: gpt-oss-safeguard-20bOutput$0.3per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o1Cached Input$7.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o1Input$15per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o1Output$60per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o1-proInput$150per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o1-proOutput$600per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3Input$2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3Output$8per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 (batch)Output$4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 MiniCached Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 MiniInput$1.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 MiniOutput$4.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini (batch)Cached Input$0.275per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini (batch)Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini (batch)Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini HighCached Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini HighInput$1.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 Mini HighOutput$4.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 ProInput$20per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o3 ProOutput$80per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 MiniCached Input$0.275per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 MiniInput$1.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 MiniOutput$4.4per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini (batch)Cached Input$0.1375per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini (batch)Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini (batch)Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini HighCached Input$0.275per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini HighInput$1.1per 1M tokensStandardSee sourceOfficial source ↗
OpenAI: o4 Mini HighOutput$4.4per 1M tokensStandardSee sourceOfficial source ↗
Perceptron: Perceptron Mk1Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Perceptron: Perceptron Mk1Output$1.5per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: SonarInput$1per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: SonarOutput$1per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Deep ResearchInput$2per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Deep ResearchOutput$8per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar ProInput$3per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar ProOutput$15per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Pro SearchInput$3per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Pro SearchOutput$15per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Reasoning ProInput$2per 1M tokensStandardSee sourceOfficial source ↗
Perplexity: Sonar Reasoning ProOutput$8per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna S 2.1Cached Input$0.009per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna S 2.1Input$0.09per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna S 2.1Output$0.18per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna S 2.1 (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna S 2.1 (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna XS 2.1Cached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna XS 2.1Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna XS 2.1Output$0.12per 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna XS 2.1 (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Poolside: Laguna XS 2.1 (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen Plus 0728Input$0.26per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen Plus 0728Output$0.78per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen-PlusCached Input$0.052per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen-PlusInput$0.26per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen-PlusOutput$0.78per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen2.5 7B InstructInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen2.5 7B InstructOutput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen2.5 VL 72B InstructCached Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen2.5 VL 72B InstructInput$0.8per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen2.5 VL 72B InstructOutput$1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 14BInput$0.2275per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 14BOutput$0.91per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22BInput$0.455per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22BOutput$1.82per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22B Instruct 2507Cached Input$0.0175per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22B Instruct 2507Input$0.0875per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22B Instruct 2507Output$0.35per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22B Thinking 2507Input$0.23per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 235B A22B Thinking 2507Output$2.3per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3BInput$0.12per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3BOutput$0.5per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3B Instruct 2507Input$0.09per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3B Instruct 2507Output$0.3per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3B Thinking 2507Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 30B A3B Thinking 2507Output$2.4per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 32BInput$0.08per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 32BOutput$0.28per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 8BInput$0.117per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 8BOutput$0.455per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder 30B A3B InstructInput$0.07per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder 30B A3B InstructOutput$0.28per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder 480B A35BCached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder 480B A35BInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder 480B A35BOutput$1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder FlashCached Input$0.039per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder FlashInput$0.195per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder FlashOutput$0.975per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder NextCached Input$0.07per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder NextInput$0.12per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder NextOutput$0.8per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder PlusCached Input$0.13per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder PlusInput$0.65per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Coder PlusOutput$3.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 MaxCached Input$0.156per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 MaxInput$0.78per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 MaxOutput$3.9per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Max ThinkingInput$0.78per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Max ThinkingOutput$3.9per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Next 80B A3B InstructInput$0.09per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Next 80B A3B InstructOutput$1.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Next 80B A3B ThinkingInput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 Next 80B A3B ThinkingOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 235B A22B InstructCached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 235B A22B InstructInput$0.21per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 235B A22B InstructOutput$1.9per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 235B A22B ThinkingInput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 235B A22B ThinkingOutput$4per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 30B A3B InstructInput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 30B A3B InstructOutput$0.6per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 30B A3B ThinkingInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 30B A3B ThinkingOutput$2.4per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 32B InstructInput$0.104per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 32B InstructOutput$0.416per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 8B InstructInput$0.117per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 8B InstructOutput$0.455per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 8B ThinkingInput$0.18per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3 VL 8B ThinkingOutput$2.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 397B A17BCached Input$0.225per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 397B A17BInput$0.55per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 397B A17BOutput$3.5per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 Plus 2026-02-15Input$0.26per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 Plus 2026-02-15Output$1.56per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 Plus 2026-04-20Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5 Plus 2026-04-20Output$1.8per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-122B-A10BInput$0.26per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-122B-A10BOutput$2.08per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-27BInput$0.195per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-27BOutput$1.56per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-35B-A3BCached Input$0.1563per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-35B-A3BInput$0.3125per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-35B-A3BOutput$1.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-9BInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-9BOutput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-9B (batch)Input$0.17per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-9B (batch)Output$0.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-FlashInput$0.065per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.5-FlashOutput$0.26per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 27BCached Input$0.03per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 27BInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 27BOutput$2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 35B A3BCached Input$0.05per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 35B A3BInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 35B A3BOutput$0.9per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 FlashInput$0.1875per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 FlashOutput$1.13per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 Max PreviewInput$1.03per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 Max PreviewOutput$6.16per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 PlusInput$0.325per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.6 PlusOutput$1.95per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 FlashCached Input$0.006per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 FlashInput$0.03per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 FlashOutput$0.13per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 MaxCached Input$0.295per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 MaxInput$1.48per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 MaxOutput$4.43per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 PlusCached Input$0.064per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 PlusInput$0.32per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.7 PlusOutput$1.28per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95BCached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95BInput$2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95BOutput$6per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95B (batch)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95B (batch)Input$2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 2.4T A95B (batch)Output$6per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 27BCached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 27BInput$0.214per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 27BOutput$2.55per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 FlashCached Input$0.016per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 FlashInput$0.15per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 FlashOutput$0.47per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 Max (0902)Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 Max (0902)Input$2per 1M tokensStandardSee sourceOfficial source ↗
Qwen: Qwen3.8 Max (0902)Output$6per 1M tokensStandardSee sourceOfficial source ↗
Qwen2.5 72B InstructInput$0.36per 1M tokensStandardSee sourceOfficial source ↗
Qwen2.5 72B InstructOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Qwen2.5 Coder 32B InstructInput$0.66per 1M tokensStandardSee sourceOfficial source ↗
Qwen2.5 Coder 32B InstructOutput$1per 1M tokensStandardSee sourceOfficial source ↗
Reka EdgeInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Reka EdgeOutput$0.1per 1M tokensStandardSee sourceOfficial source ↗
Reka Flash 3Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Reka Flash 3Output$0.2per 1M tokensStandardSee sourceOfficial source ↗
Relace: Relace Apply 3Input$0.85per 1M tokensStandardSee sourceOfficial source ↗
Relace: Relace Apply 3Output$1.25per 1M tokensStandardSee sourceOfficial source ↗
Relace: Relace SearchInput$1per 1M tokensStandardSee sourceOfficial source ↗
Relace: Relace SearchOutput$3per 1M tokensStandardSee sourceOfficial source ↗
ReMM SLERP 13BInput$0.35per 1M tokensStandardSee sourceOfficial source ↗
ReMM SLERP 13BOutput$0.65per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu MaxCached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu MaxInput$2per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu MaxOutput$6per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu UltraCached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu UltraInput$5per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu UltraOutput$30per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu Ultra v2Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu Ultra v2Input$5per 1M tokensStandardSee sourceOfficial source ↗
Sakana: Fugu Ultra v2Output$30per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3 8B LunarisInput$0.04per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3 8B LunarisOutput$0.05per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3.1 Euryale 70B v2.2Input$0.85per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3.1 Euryale 70B v2.2Output$0.85per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3.3 Euryale 70BInput$0.65per 1M tokensStandardSee sourceOfficial source ↗
Sao10K: Llama 3.3 Euryale 70BOutput$0.75per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20 Multi-AgentCached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20 Multi-AgentInput$1.25per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.20 Multi-AgentOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3Input$1.25per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3Output$2.5per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3 (batch)Cached Input$0.16per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3 (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.3 (batch)Output$2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.5Cached Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.5Input$2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.5Output$6per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.6Cached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.6Input$2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok 4.6Output$6per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok Build 0.1Cached Input$0.2per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok Build 0.1Input$1per 1M tokensStandardSee sourceOfficial source ↗
SpaceXAI: Grok Build 0.1Output$2per 1M tokensStandardSee sourceOfficial source ↗
StepFun: Step 3.5 FlashInput$0.1per 1M tokensStandardSee sourceOfficial source ↗
StepFun: Step 3.5 FlashOutput$0.3per 1M tokensStandardSee sourceOfficial source ↗
StepFun: Step 3.7 FlashCached Input$0.04per 1M tokensStandardSee sourceOfficial source ↗
StepFun: Step 3.7 FlashInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
StepFun: Step 3.7 FlashOutput$1.15per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hunyuan A13B InstructInput$0.14per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hunyuan A13B InstructOutput$0.57per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-1.8BInput$0.044per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-1.8BOutput$0.177per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-30B-A3BInput$0.074per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-30B-A3BOutput$0.295per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-7BInput$0.074per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy-MT2-7BOutput$0.295per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3Cached Input$0.033per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3Input$0.132per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3Output$0.528per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3 previewCached Input$0.06per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3 previewInput$0.18per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy3 previewOutput$0.6per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy4 previewCached Input$0.042per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy4 previewInput$0.834per 1M tokensStandardSee sourceOfficial source ↗
Tencent: Hy4 previewOutput$2.5per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Cydonia 24B V4.1Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Cydonia 24B V4.1Input$0.3per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Cydonia 24B V4.1Output$0.5per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Skyfall 36B V2Cached Input$0.25per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Skyfall 36B V2Input$0.55per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: Skyfall 36B V2Output$0.8per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: UnslopNemo 12BInput$0.4per 1M tokensStandardSee sourceOfficial source ↗
TheDrummer: UnslopNemo 12BOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: InklingCached Input$0.17per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: InklingInput$1per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: InklingOutput$4.05per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling (batch)Cached Input$0.17per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling (batch)Input$1per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling (batch)Output$4.05per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling SmallCached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling SmallInput$0.45per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling SmallOutput$1.2per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling Small (batch)Cached Input$0.1per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling Small (batch)Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling Small (batch)Output$1.2per 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling Small (free)InputFreeper 1M tokensStandardSee sourceOfficial source ↗
Thinking Machines: Inkling Small (free)OutputFreeper 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 3Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 3Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 3Output$0.6per 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 4Cached Input$0.018per 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 4Input$0.09per 1M tokensStandardSee sourceOfficial source ↗
Upstage: Solar Pro 4Output$0.36per 1M tokensStandardSee sourceOfficial source ↗
Venice: UncensoredInput$0.2per 1M tokensStandardSee sourceOfficial source ↗
Venice: UncensoredOutput$0.9per 1M tokensStandardSee sourceOfficial source ↗
WizardLM-2 8x22BInput$0.62per 1M tokensStandardSee sourceOfficial source ↗
WizardLM-2 8x22BOutput$0.62per 1M tokensStandardSee sourceOfficial source ↗
Writer: Palmyra X5Input$0.6per 1M tokensStandardSee sourceOfficial source ↗
Writer: Palmyra X5Output$6per 1M tokensStandardSee sourceOfficial source ↗
xAI: Grok LatestCached Input$0.5per 1M tokensStandardSee sourceOfficial source ↗
xAI: Grok LatestInput$2per 1M tokensStandardSee sourceOfficial source ↗
xAI: Grok LatestOutput$6per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5Cached Input$0.0028per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5Input$0.14per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5Output$0.28per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5-ProCached Input$0.0036per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5-ProInput$0.435per 1M tokensStandardSee sourceOfficial source ↗
Xiaomi: MiMo-V2.5-ProOutput$0.87per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5Cached Input$0.11per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5Input$0.6per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5 AirCached Input$0.025per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5 AirInput$0.13per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5 AirOutput$0.85per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5VCached Input$0.11per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5VInput$0.6per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.5VOutput$1.8per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6Cached Input$0.08per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6Input$0.43per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6Output$1.75per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6VCached Input$0.055per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6VInput$0.3per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.6VOutput$0.9per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.7Cached Input$0.08per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.7Input$0.4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.7Output$1.75per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.7 FlashInput$0.0605per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 4.7 FlashOutput$0.4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5Cached Input$0.12per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5Input$0.6per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5Output$1.92per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5 TurboCached Input$0.24per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5 TurboInput$1.2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5 TurboOutput$4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.1Cached Input$0.1794per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.1Input$0.966per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.1Output$3.04per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2Cached Input$0.15per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2Input$0.6per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2Output$2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2 (batch)Cached Input$0.07per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2 (batch)Input$0.7per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.2 (batch)Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3Cached Input$0.26per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3Input$1.4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3Output$4.4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 (batch)Cached Input$0.13per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 (batch)Input$0.7per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 (batch)Output$2.2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 FlashCached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 FlashInput$0.075per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 FlashOutput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 Flash (batch)Cached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 Flash (batch)Input$0.075per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5.3 Flash (batch)Output$0.25per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5V TurboCached Input$0.24per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5V TurboInput$1.2per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM 5V TurboOutput$4per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM Flash LatestCached Input$0.015per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM Flash LatestInput$0.075per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM Flash LatestOutput$0.25per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM LatestCached Input$0.1639per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM LatestInput$0.8727per 1M tokensStandardSee sourceOfficial source ↗
Z.ai: GLM LatestOutput$3.36per 1M tokensStandardSee sourceOfficial source ↗

Last verified · Source ↗

Inspect the identifier, category and condition before comparing adjacent rows. Unknown data should remain an unresolved fact, and a non-token category should not be dropped merely because the application’s first calculator view is token-focused.

Estimate your cost

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

Preserve the workload in the OpenRouter cost calculator and record the routing assumptions beside the result.

Use the same input and output requirements when comparing routes. A response that needs a repair call or another model should be counted as a multi-call task. Otherwise the least expensive individual call can appear to win a comparison while producing a larger cost per accepted result.

Three worked scenarios

For a chatbot, decide whether continuity requires a fixed model or permits a fallback. Test the same conversation state against every allowed candidate. Include the cost of a follow-up model call if an initial answer fails validation rather than presenting the nominal first call as the whole interaction.

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

For summarization, identify whether requests repeat a stable document prefix or each contain unrelated material. Cache assumptions should follow that structure. Keep an uncached scenario available and inspect whether the actual serving route supports the cache category used in the estimate.

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

For a coding assistant, include the repository context, output patch and the checks required to accept it. A provider-routing change can affect operational behavior even when the model family remains familiar, so record the resolved route alongside usage and test results.

Interactive tool

Estimate your API costs

Your text and estimates stay in this browser. No API requests are sent to model providers.

Loading verified model records…

How to reduce cost on OpenRouter

Provider routing exposes price-oriented selection and maximum-price conditions, while prompt-caching behavior depends on the selected model and provider path. Official documentation.

Use a price constraint after specifying mandatory capabilities. A cheap endpoint that ignores a field essential to your output contract is not a useful optimization. Keep required parameter support explicit and test both ordinary and exceptional outputs before widening the routing set.

The caching guide describes provider-specific cache behavior and the usage categories that indicate reuse. Official documentation.

Measure reuse on a stable fixture and preserve the serving path while investigating differences. Avoid assuming a cache warmed through one route will produce the same result through every other endpoint. First establish the request structure and observed behavior; then use that evidence in the workload scenario.

Compare smaller candidates against a fixed acceptance criterion. If a simpler task can be routed to a less costly model, make that task boundary explicit in application logic and test the classifier or routing rule. A cost-saving policy should not quietly send difficult requests to a candidate that has never passed the relevant cases.

Billing, credits and limits notes

OpenRouter distinguishes account balance and per-key credit limits from request-rate restrictions. Official documentation.

Inspect the account setting that actually rejected a request before changing funds or traffic. A key cap, workspace budget and account balance can require different owner actions. Keep a record of which boundary was intended for the application so a quick operational fix does not silently remove its spending control.

Review free-route conditions and capacity and credit diagnostics. Do not let a fallback policy turn a deliberately free experiment into an unreviewed paid workload.

Price change history

  1. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: cached_input · USD: 0.062364500 · Unit: per 1M tokens → Price category: cached_input · USD: 0.0615235 · Unit: per 1M tokensSource ↗
  2. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: output · USD: 1.496748000 · Unit: per 1M tokens → Price category: output · USD: 1.476564 · Unit: per 1M tokensSource ↗
  3. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: input · USD: 0.748374000 · Unit: per 1M tokens → Price category: input · USD: 0.738282 · Unit: per 1M tokensSource ↗
  4. OpenRouter · DeepSeek: DeepSeek V4 Flash 0423 — Amount UsdPrice category: cached_input · USD: 0.013244000 · Unit: per 1M tokens → Price category: cached_input · USD: 0.013188 · Unit: per 1M tokensSource ↗
  5. OpenRouter · DeepSeek: DeepSeek V4 Flash 0423 — Amount UsdPrice category: output · USD: 0.132440000 · Unit: per 1M tokens → Price category: output · USD: 0.13188 · Unit: per 1M tokensSource ↗
  6. OpenRouter · DeepSeek: DeepSeek V4 Flash 0423 — Amount UsdPrice category: input · USD: 0.066220000 · Unit: per 1M tokens → Price category: input · USD: 0.06594 · Unit: per 1M tokensSource ↗
  7. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: cached_input · USD: 0.063220000 · Unit: per 1M tokens → Price category: cached_input · USD: 0.0623645 · Unit: per 1M tokensSource ↗
  8. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: output · USD: 1.517280000 · Unit: per 1M tokens → Price category: output · USD: 1.496748 · Unit: per 1M tokensSource ↗
  9. OpenRouter · DeepSeek: DeepSeek V4 Pro 0423 — Amount UsdPrice category: input · USD: 0.758640000 · Unit: per 1M tokens → Price category: input · USD: 0.748374 · Unit: per 1M tokensSource ↗
  10. OpenRouter · DeepSeek: DeepSeek V4 Pro 0813 — Amount UsdPrice category: cached_input · USD: 0.018396000 · Unit: per 1M tokens → Price category: cached_input · USD: 0.019316 · Unit: per 1M tokensSource ↗

Subscribe to the changelog RSS feed

Recalculate the previous request assumptions when a route price changes. Keep the rate effect separate from a different resolved model, a changed cache share or a longer generated response.

Last verified · Source ↗

Frequently asked questions

Can every catalog rate be compared as text tokens?
No. Retain each billable unit and qualifying condition.
Does BYOK remove the need to inspect OpenRouter billing rules?
No. Review the current BYOK fee and fallback conditions along with the upstream provider account.
Can a price ceiling make a request fail?
It can reduce the set of eligible routes; define the application behavior when none remains.
Should I assume caches work identically across routes?
Use the provider-specific behavior and measured usage for the actual serving path.
What should a fallback cost forecast include?
The allowed candidate routes and additional calls needed to reach an accepted task.
Does a low input rate prove lowest application cost?
No. Output, repair work and other applicable categories can change the total.

Sources

Last verified · Source ↗