gin

Gin / Resources

LLM cost calculator

What an AI feature costs per call, per day and per month, on the model you use today and on the cheaper ones that might do the same job. 448 models, priced from OpenRouter.

Prices per 1M tokens · updated October 8, 2026

Cost calculator

Typical calls
Monthly cost (30 days) on your model and on the newest models from each major provider
ModelPer callPer dayPer monthvs yours
Loading prices…
Same price for a model doesn’t mean same quality. Gin grades cheaper models on your calls before it routes any. Grade my calls for free →

How the math works

Every provider bills per token, with separate prices for what you send and what comes back:

cost per call = input tokens × input price
              + output tokens × output price
              − cached input tokens × (input price − cache-read price)

Prices are quoted per million tokens. For example, a call to Claude Sonnet 5.5 with 3,000 input and 500 output tokens costs 3,000 × $2.00/1M + 500 × $10.00/1M. Multiply by calls per day and by 30 for the month.

  • A token is about four characters of English, or three quarters of a word. Code, JSON and other languages take more.
  • Output costs more than input, usually 3–8×, because it is generated one token at a time. Long answers, reasoning tokens and verbose JSON add up fast.
  • Cached input is cheap. When the start of your prompt repeats, most providers bill it at 10–50% of the input price. See how prompt caching is priced.
  • Reasoning tokens are output tokens. A model that thinks for 2,000 tokens before a 200-token answer bills 2,200 output tokens.

Where to get your real numbers

The calculator is only as good as its inputs. Your provider’s usage page shows tokens by model and by key, but not by feature, so “what does the summarizer cost?” stays a guess. Two ways to get it:

  1. Log usage.prompt_tokens, usage.completion_tokens and cached tokens from every response, tagged with the feature that made the call.
  2. Or wrap your client with Gin in one line, fetch: gin({ useCase: 'summarize' }), and get spend by use case in a free daily email, with cheaper models already graded on your calls.

Token prices for every model on OpenRouter

448 models · $ per 1M tokens · from the OpenRouter API, October 8, 2026
ModelInputOutputCache readContext
Claude Fable Latest~anthropic/claude-fable-latest$10.00$50.00$0.251M
Claude Haiku Latest~anthropic/claude-haiku-latest$0.10$0.50$0.011M
Claude Sonnet Latest~anthropic/claude-sonnet-latest$2.00$10.00$0.101M
Claude Opus Latest~anthropic/claude-opus-latest$4.00$20.00$0.201M
DeepSeek Pro Latest~deepseek/deepseek-pro-latest$0.135$8.00$0.101M
DeepSeek Flash Latest~deepseek/deepseek-flash-latest$0.0030$2.40$0.00301M
DeepSeek V4 Flash Latest~deepseek/deepseek-v4-flash-latest$0.01$1.28$0.011M
Gemini Pro Latest~google/gemini-pro-latest$2.00$12.00$0.201M
Gemini Flash Latest~google/gemini-flash-latest$0.75$3.75$0.0751M
Kimi Latest~moonshotai/kimi-latest$0.56$13.00$0.451M
GPT Astra Latest~openai/gpt-astra-latest$10.00$50.00$1.001.1M
GPT Sol Latest~openai/gpt-sol-latest$2.00$10.00$0.101.1M
GPT Terra Latest~openai/gpt-terra-latest$2.00$12.00$0.201.1M
GPT Luna Latest~openai/gpt-luna-latest$0.10$0.50$0.011.1M
GPT Mini Latest~openai/gpt-mini-latest$0.75$4.50$0.075400K
Grok Latest~x-ai/grok-latest$2.00$6.00$0.50500K
GLM Flash Latest~z-ai/glm-flash-latest$0.032$1.41$0.021M
GLM Latest~z-ai/glm-latest$0.036$12.00$0.0341M
Aion 3.5 Miniaion-labs/aion-3.5-mini$0.70$1.40$0.18262K
Aion 3.5aion-labs/aion-3.5$3.00$6.00$0.75262K
Aion-3.0-Miniaion-labs/aion-3.0-mini$0.70$1.40$0.18131K
Aion-3.0aion-labs/aion-3.0$3.00$6.00$0.75131K
Aion-2.0aion-labs/aion-2.0$0.80$1.60$0.20131K
Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b$0.80$1.60—33K
Nova 2 Liteamazon/nova-2-lite-v1$0.30$2.50—1M
Nova Premier 1.0amazon/nova-premier-v1$2.50$12.50$0.6251M
Nova Lite 1.0amazon/nova-lite-v1$0.06$0.24—300K
Nova Micro 1.0amazon/nova-micro-v1$0.035$0.14—128K
Nova Pro 1.0amazon/nova-pro-v1$0.80$3.20—300K
Magnum v4 72Banthracite-org/magnum-v4-72b$2.50$5.00—33K
Claude Haiku 5.5anthropic/claude-haiku-5.5$0.10$0.50$0.011M
Claude Haiku 5.5 (batch)anthropic/claude-haiku-5.5:batch$0.05$0.25$0.00501M
Claude Sonnet 5.5anthropic/claude-sonnet-5.5$2.00$10.00$0.101M
Claude Sonnet 5.5 (batch)anthropic/claude-sonnet-5.5:batch$1.00$5.00$0.051M
Claude Opus 5.5anthropic/claude-opus-5.5$4.00$20.00$0.201M
Claude Opus 5.5 (batch)anthropic/claude-opus-5.5:batch$2.00$10.00$0.101M
Claude Fable 5.1anthropic/claude-fable-5.1$10.00$50.00$0.251M
Claude Fable 5.1 (batch)anthropic/claude-fable-5.1:batch$5.00$25.00$0.1251M
Claude Opus 5anthropic/claude-opus-5$5.00$25.00$0.501M
Claude Opus 5 (batch)anthropic/claude-opus-5:batch$2.50$12.50$0.251M
Claude Sonnet 5anthropic/claude-sonnet-5$2.00$10.00$0.201M
Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch$1.00$5.00$0.101M
Claude Fable 5anthropic/claude-fable-5$10.00$50.00$1.001M
Claude Fable 5 (batch)anthropic/claude-fable-5:batch$5.00$25.00$0.501M
Claude Opus 4.8anthropic/claude-opus-4.8$5.00$25.00$0.501M
Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch$2.50$12.50$0.251M
Claude Opus 4.7anthropic/claude-opus-4.7$5.00$25.00$0.501M
Claude Opus 4.7 (batch)anthropic/claude-opus-4.7:batch$2.50$12.50$0.251M
Claude Sonnet 4.6anthropic/claude-sonnet-4.6$3.00$15.00$0.301M
Claude Sonnet 4.6 (batch)anthropic/claude-sonnet-4.6:batch$1.50$7.50$0.151M
Claude Opus 4.6anthropic/claude-opus-4.6$5.00$25.00$0.501M
Claude Opus 4.6 (batch)anthropic/claude-opus-4.6:batch$2.50$12.50$0.251M
Claude Opus 4.5anthropic/claude-opus-4.5$5.00$25.00$0.50200K
Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch$2.50$12.50$0.25200K
Claude Haiku 4.5anthropic/claude-haiku-4.5$1.00$5.00$0.10200K
Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch$0.50$2.50$0.05200K
Claude Sonnet 4.5anthropic/claude-sonnet-4.5$3.00$15.00$0.301M
Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch$1.50$7.50$0.151M
Claude Opus 4.1anthropic/claude-opus-4.1$15.00$75.00$1.50200K
Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch$7.50$37.50$0.75200K
Claude Sonnet 4anthropic/claude-sonnet-4$3.00$15.00$0.30200K
Apodex 1.1 Mini (free)apodex/apodex-1.1-mini:free$0$0—262K
Trinity Large Thinkingarcee-ai/trinity-large-thinking$0.25$0.80$0.06262K
ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b$0.42$1.25—123K
UI-TARS 7B bytedance/ui-tars-1.5-7b$0.10$0.20$0.10128K
Seed 2.1 Turbobytedance-seed/seed-2-1-turbo$0.50$2.50—262K
Seed-2.0-Codebytedance-seed/seed-2.0-code$0.50$3.00—262K
Seed-2.0-Litebytedance-seed/seed-2.0-lite$0.25$2.00—262K
Seed-2.0-Minibytedance-seed/seed-2.0-mini$0.10$0.40—262K
Seed 1.6 Flashbytedance-seed/seed-1.6-flash$0.075$0.30—262K
Seed 1.6bytedance-seed/seed-1.6$0.25$2.00—262K
Uncensoredcognitivecomputations/dolphin-mistral-24b-venice-edition$0.20$0.90—128K
Command A+cohere/command-a-plus$0.30$1.50$0.15192K
North Mini Code (free)cohere/north-mini-code:free$0$0—256K
Command Acohere/command-a$2.50$10.00—256K
Command R7B (12-2024)cohere/command-r7b-12-2024$0.037$0.15—128K
Command R (08-2024)cohere/command-r-08-2024$0.15$0.60—128K
Command R+ (08-2024)cohere/command-r-plus-08-2024$2.50$10.00—128K
DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash$0.30$1.20$0.00601M
DeepSeek V4.1 Flash (batch)deepseek/deepseek-v4.1-flash:batch$0.112$0.336$0.00341M
DeepSeek V4 Flash Vision Expdeepseek/deepseek-v4-flash-vision-exp$0.216$0.647$0.00691M
DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813$1.32$3.96$0.0441M
DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731$0.01$1.28$0.011M
DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro$0.955$1.91$0.081M
DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash$0.0097$1.28$0.00971M
DeepSeek V3.2deepseek/deepseek-v3.2$0.259$0.42$0.135164K
DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp$0.27$0.41—164K
DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus$0.27$1.00—164K
DeepSeek V3.1deepseek/deepseek-chat-v3.1$0.25$0.95$0.13164K
R1 0528deepseek/deepseek-r1-0528$0.50$2.15$0.35164K
DeepSeek V3 0324deepseek/deepseek-chat-v3-0324$0.29$1.14$0.11164K
R1deepseek/deepseek-r1$0.70$2.50—64K
DeepSeek V3deepseek/deepseek-chat$0.257$1.03—164K
Dots3-Note Preview (free)dots-studio/dots-3-note-preview:free$0$0—512K
Ember-1fireworks/ember-1$3.00$15.00$0.301M
Gemini 3.8 Flashgoogle/gemini-3.8-flash$0.75$3.75$0.0751M
Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch$0.375$1.88$0.0371M
Gemini 3.7 Flashgoogle/gemini-3.7-flash$0.75$3.75$0.0751M
Gemini 3.7 Flash (batch)google/gemini-3.7-flash:batch$0.375$1.88$0.0371M
Gemini 3.6 Flashgoogle/gemini-3.6-flash$0.75$3.75$0.0751M
Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch$0.375$1.88$0.0371M
Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite$0.30$2.50$0.031M
Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch$0.15$1.25$0.0151M
Gemini 3.5 Flashgoogle/gemini-3.5-flash$1.50$9.00$0.151M
Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch$0.75$4.50$0.0751M
Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite$0.25$1.50$0.0251M
Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch$0.125$0.75$0.0131M
Gemma 4 26B A4B google/gemma-4-26b-a4b-it$0.09$0.30$0.05262K
Gemma 4 26B A4B (free)google/gemma-4-26b-a4b-it:free$0$0—262K
Gemma 4 31Bgoogle/gemma-4-31b-it$0.09$0.34$0.05262K
Gemma 4 31B (free)google/gemma-4-31b-it:free$0$0—262K
Lyria 3 Pro Previewgoogle/lyria-3-pro-preview$0$0—1M
Lyria 3 Clip Previewgoogle/lyria-3-clip-preview$0$0—1M
Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview$0.25$1.50$0.0251M
Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools$2.00$12.00$0.201M
Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview$2.00$12.00$0.201M
Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch$1.00$6.00—1M
Gemini 3 Flash Previewgoogle/gemini-3-flash-preview$0.50$3.00$0.051M
Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch$0.25$1.50—1M
Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite$0.10$0.40$0.011M
Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch$0.05$0.20$0.011M
Gemini 2.5 Flashgoogle/gemini-2.5-flash$0.30$2.50$0.031M
Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch$0.15$1.25$0.031M
Gemini 2.5 Progoogle/gemini-2.5-pro$1.25$10.00$0.1251M
Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch$0.625$5.00$0.1251M
Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview$1.25$10.00$0.1251M
Gemma 3 4Bgoogle/gemma-3-4b-it$0.05$0.10—131K
Gemma 3 12Bgoogle/gemma-3-12b-it$0.05$0.15—131K
Gemma 3 27Bgoogle/gemma-3-27b-it$0.08$0.45$0.04131K
Gemma 2 27Bgoogle/gemma-2-27b-it$0.65$0.65—8K
MythoMax 13Bgryphe/mythomax-l2-13b$0.08$0.11—8K
Granite 4.2 8Bibm-granite/granite-4.2-8b$0.06$0.25$0.015131K
Granite 4.0 Microibm-granite/granite-4.0-h-micro$0.017$0.112—131K
Mercury 2.5inception/mercury-2.5$0.04$0.15$0.0040260K
Mercury 2inception/mercury-2$0.25$0.75$0.025128K
Ling 3.1 Flashinclusionai/ling-3.1-flash$0$0—262K
Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl$0.021$0.062$0.0042262K
Ling 3.0 Flash Sante (free)inclusionai/ling-3.0-flash-sante:free$0$0—262K
Ling 3.0 Flash Fininclusionai/ling-3.0-flash-fin$0.042$0.123$0.0084262K
Ling 3.0 Flashinclusionai/ling-3.0-flash$0.021$0.063$0.0042262K
Schematron V2 Turboinference-net/schematron-v2-turbo$0.03$0.15$0.03128K
Schematron V2 Smallinference-net/schematron-v2-small$0.05$0.23$0.05128K
LFM2.5-2.6B (free)liquid/lfm-2.5-2.6b:free$0$0—66K
Weaver (alpha)mancer/weaver$0.40$0.75—8K
LongCat 2.0meituan/longcat-2.0$0.30$1.20$0.00601M
Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor$0.10$0.20$0.00201M
Muse Spark 1.3meta/muse-spark-1.3$1.25$4.25$0.151M
Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor$0.10$0.20$0.00201M
Muse Glimmer 30Bmeta/muse-glimmer-30b$0.30$1.20$0.04131K
Muse Spark 1.2meta/muse-spark-1.2$1.25$4.25$0.151M
Muse Spark 1.1meta/muse-spark-1.1$1.25$4.25$0.151M
Llama Guard 4 12Bmeta-llama/llama-guard-4-12b$0.18$0.18—164K
Llama 4 Maverickmeta-llama/llama-4-maverick$0.188$0.652$0.051M
Llama 4 Scoutmeta-llama/llama-4-scout$0.10$0.30—1.3M
Llama 3.3 70B Instructmeta-llama/llama-3.3-70b-instruct$0.10$0.32—131K
Llama 3.2 1B Instructmeta-llama/llama-3.2-1b-instruct$0.027$0.201—60K
Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct$0.05$0.33—131K
Llama 3.1 70B Instructmeta-llama/llama-3.1-70b-instruct$0.40$0.40—131K
Llama 3.1 8B Instructmeta-llama/llama-3.1-8b-instruct$0.05$0.08$0.025131K
Phi 4microsoft/phi-4$0.07$0.14—16K
WizardLM-2 8x22Bmicrosoft/wizardlm-2-8x22b$0.62$0.62—66K
MiniMax M3minimax/minimax-m3$0.30$1.20$0.061M
MiniMax M2.7minimax/minimax-m2.7$0.21$0.84$0.042205K
MiniMax M2.5minimax/minimax-m2.5$0.27$1.08$0.027205K
MiniMax M2-herminimax/minimax-m2-her$0.30$1.20$0.0366K
MiniMax M2.1minimax/minimax-m2.1$0.30$1.20$0.03205K
MiniMax M2minimax/minimax-m2$0.30$1.20—205K
MiniMax M1minimax/minimax-m1$0.55$2.20—1M
MiniMax-01minimax/minimax-01$0.20$1.10—1M
Mistral Large 4mistralai/mistral-large-4-0$0.68$2.09$0.07524K
Mistral Medium 3.5mistralai/mistral-medium-3-5$1.50$7.50—262K
Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch$0.75$3.75—262K
Mistral Small 4mistralai/mistral-small-2603$0.15$0.60$0.015262K
Mistral Small 4 (batch)mistralai/mistral-small-2603:batch$0.075$0.30$0.0075262K
Devstral 2 2512mistralai/devstral-2512$0.40$2.00$0.04262K
Ministral 3 14B 2512mistralai/ministral-14b-2512$0.20$0.20$0.02262K
Ministral 3 8B 2512mistralai/ministral-8b-2512$0.15$0.15$0.015262K
Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch$0.075$0.075$0.0075262K
Ministral 3 3B 2512mistralai/ministral-3b-2512$0.10$0.10$0.01131K
Mistral Large 3 2512mistralai/mistral-large-2512$0.50$1.50$0.05262K
Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch$0.25$0.75$0.025262K
Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507$0.10$0.30$0.0133K
Mistral Medium 3.1mistralai/mistral-medium-3.1$0.40$2.00$0.04131K
Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch$0.20$1.00$0.02131K
Codestral 2508mistralai/codestral-2508$0.30$0.90$0.03256K
Codestral 2508 (batch)mistralai/codestral-2508:batch$0.15$0.45$0.015256K
Mistral Small 3.2 24Bmistralai/mistral-small-3.2-24b-instruct$0.094$0.25—256K
Mistral Medium 3mistralai/mistral-medium-3$0.40$2.00$0.04131K
Mistral Small 3.1 24Bmistralai/mistral-small-3.1-24b-instruct$0.351$0.555—128K
Sabamistralai/mistral-saba$0.20$0.60$0.0233K
Mistral Small 3mistralai/mistral-small-24b-instruct-2501$0.05$0.08—33K
Mistral Large 2407mistralai/mistral-large-2407$2.00$6.00$0.20131K
Mistral Nemomistralai/mistral-nemo$0.019$0.03—131K
Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct$2.00$6.00$0.2066K
Mistral Largemistralai/mistral-large$2.00$6.00$0.20128K
Kimi K3moonshotai/kimi-k3$0.58$12.30$0.401M
Kimi K3 (batch)moonshotai/kimi-k3:batch$2.28$11.40$0.2281M
Kimi K2.7 Codemoonshotai/kimi-k2.7-code$0.671$3.35$0.18262K
Kimi K2.6moonshotai/kimi-k2.6$0.465$2.45$0.098262K
Kimi K2.5moonshotai/kimi-k2.5$0.45$2.25$0.07262K
Kimi K2 Thinkingmoonshotai/kimi-k2-thinking$0.60$2.50—262K
Kimi K2 0905moonshotai/kimi-k2-0905$0.60$2.50—262K
Kimi K2 0711moonshotai/kimi-k2$0.57$2.30—131K
Morph V3 Largemorph/morph-v3-large$0.90$1.90—262K
Morph V3 Fastmorph/morph-v3-fast$0.80$1.20—82K
Nex-N2.5-Mininex-agi/nex-n2.5-mini$0.025$0.10$0.0025262K
Nex-N2.5-Pronex-agi/nex-n2.5-pro$0.075$0.25$0.015262K
Hermes 4 405Bnousresearch/hermes-4-405b$1.00$3.00—131K
Hermes 3 70B Instructnousresearch/hermes-3-llama-3.1-70b$0.70$0.70—131K
Hermes 3 405B Instructnousresearch/hermes-3-llama-3.1-405b$1.00$1.00—131K
Nemotron 3.5 Lightningnvidia/nemotron-3.5-lightning$0.049$0.14$0.025262K
Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free$0$0—1M
Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety$0.20$0.20—131K
Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free$0$0—128K
Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b$0.50$2.20$0.10262K
Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free$0$0—1M
Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free$0$0—256K
Nemotron 3 Supernvidia/nemotron-3-super-120b-a12b$0.085$0.40—262K
Nemotron 3 Super (free)nvidia/nemotron-3-super-120b-a12b:free$0$0—262K
Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b$0.06$0.24—262K
GPT-6.1 Sol Proopenai/gpt-6.1-sol-pro$2.00$10.00$0.101.1M
GPT-6.1 Solopenai/gpt-6.1-sol$2.00$10.00$0.101.1M
GPT-6 Luna Proopenai/gpt-6-luna-pro$0.10$0.50$0.011.1M
GPT-6 Luna Pro (batch)openai/gpt-6-luna-pro:batch$0.05$0.25$0.00501.1M
GPT-6 Lunaopenai/gpt-6-luna$0.10$0.50$0.011.1M
GPT-6 Luna (batch)openai/gpt-6-luna:batch$0.05$0.25$0.00501.1M
GPT-6 Sol Proopenai/gpt-6-sol-pro$2.00$10.00$0.201.1M
GPT-6 Sol Pro (batch)openai/gpt-6-sol-pro:batch$1.00$5.00$0.101.1M
GPT-6 Solopenai/gpt-6-sol$2.00$10.00$0.201.1M
GPT-6 Sol (batch)openai/gpt-6-sol:batch$1.00$5.00$0.101.1M
GPT-6 Astraopenai/gpt-6-astra$10.00$50.00$1.001.1M
GPT-6 Astra (batch)openai/gpt-6-astra:batch$5.00$25.00$0.501.1M
GPT-6 Astra Proopenai/gpt-6-astra-pro$10.00$50.00$1.001.1M
GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch$5.00$25.00$0.501.1M
GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro$0.20$1.20$0.021.1M
GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batch$0.10$0.60$0.011.1M
GPT-5.6 Lunaopenai/gpt-5.6-luna$0.20$1.20$0.021.1M
GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch$0.10$0.60$0.011.1M
GPT-5.6 Terra Proopenai/gpt-5.6-terra-pro$2.00$12.00$0.201.1M
GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch$1.00$6.00$0.101.1M
GPT-5.6 Terraopenai/gpt-5.6-terra$2.00$12.00$0.201.1M
GPT-5.6 Terra (batch)openai/gpt-5.6-terra:batch$1.00$6.00$0.101.1M
GPT-5.6 Sol Proopenai/gpt-5.6-sol-pro$2.00$10.00$0.201.1M
GPT-5.6 Sol Pro (batch)openai/gpt-5.6-sol-pro:batch$1.00$5.00$0.101.1M
GPT-5.6 Solopenai/gpt-5.6-sol$2.00$10.00$0.201.1M
GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch$1.00$5.00$0.101.1M
GPT Chat Latestopenai/gpt-chat-latest$5.00$30.00$0.50400K
GPT-5.5 Proopenai/gpt-5.5-pro$30.00$180.00—1.1M
GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch$15.00$90.00—1.1M
GPT-5.5openai/gpt-5.5$5.00$30.00$0.501.1M
GPT-5.5 (batch)openai/gpt-5.5:batch$2.50$15.00$0.251.1M
GPT-5.4 Nanoopenai/gpt-5.4-nano$0.20$1.25$0.02400K
GPT-5.4 Nano (batch)openai/gpt-5.4-nano:batch$0.10$0.625$0.01400K
GPT-5.4 Miniopenai/gpt-5.4-mini$0.75$4.50$0.075400K
GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch$0.375$2.25$0.037400K
GPT-5.4 Proopenai/gpt-5.4-pro$30.00$180.00—1.1M
GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch$15.00$90.00—1.1M
GPT-5.4openai/gpt-5.4$2.50$15.00$0.251.1M
GPT-5.4 (batch)openai/gpt-5.4:batch$1.25$7.50$0.1251.1M
GPT-5.3-Codexopenai/gpt-5.3-codex$1.75$14.00$0.175400K
GPT-5.2-Codexopenai/gpt-5.2-codex$1.75$14.00$0.175400K
GPT-5.2 Chatopenai/gpt-5.2-chat$1.75$14.00$0.175128K
GPT-5.2 Proopenai/gpt-5.2-pro$21.00$168.00—400K
GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch$10.50$84.00—400K
GPT-5.2openai/gpt-5.2$1.75$14.00$0.175400K
GPT-5.2 (batch)openai/gpt-5.2:batch$0.875$7.00$0.087400K
GPT-5.1-Codex-Maxopenai/gpt-5.1-codex-max$1.25$10.00$0.125400K
GPT-5.1openai/gpt-5.1$1.25$10.00$0.125400K
GPT-5.1 (batch)openai/gpt-5.1:batch$0.625$5.00$0.063400K
GPT-5.1-Codexopenai/gpt-5.1-codex$1.25$10.00$0.13400K
GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini$0.25$2.00$0.03400K
gpt-oss-safeguard-20bopenai/gpt-oss-safeguard-20b$0.075$0.30$0.037131K
GPT-5 Proopenai/gpt-5-pro$15.00$120.00—400K
GPT-5 Pro (batch)openai/gpt-5-pro:batch$7.50$60.00—400K
GPT-5openai/gpt-5$1.25$10.00$0.125400K
GPT-5 (batch)openai/gpt-5:batch$0.625$5.00$0.063400K
GPT-5 Miniopenai/gpt-5-mini$0.25$2.00$0.025400K
GPT-5 Mini (batch)openai/gpt-5-mini:batch$0.125$1.00$0.013400K
GPT-5 Nanoopenai/gpt-5-nano$0.05$0.40$0.0050400K
GPT-5 Nano (batch)openai/gpt-5-nano:batch$0.025$0.20$0.0025400K
gpt-oss-120bopenai/gpt-oss-120b$0.037$0.17—131K
gpt-oss-120b (batch)openai/gpt-oss-120b:batch$0.03$0.136—131K
gpt-oss-20bopenai/gpt-oss-20b$0.018$0.09$0.0090131K
gpt-oss-20b (batch)openai/gpt-oss-20b:batch$0.024$0.112—131K
o3 Proopenai/o3-pro$20.00$80.00—200K
o4 Mini Highopenai/o4-mini-high$1.10$4.40$0.275200K
o3openai/o3$2.00$8.00$0.50200K
o3 (batch)openai/o3:batch$1.00$4.00$0.25200K
o4 Miniopenai/o4-mini$1.10$4.40$0.275200K
o4 Mini (batch)openai/o4-mini:batch$0.55$2.20$0.138200K
GPT-4.1openai/gpt-4.1$2.00$8.00$0.501M
GPT-4.1 (batch)openai/gpt-4.1:batch$1.00$4.00$0.251M
GPT-4.1 Miniopenai/gpt-4.1-mini$0.40$1.60$0.101M
GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch$0.20$0.80$0.051M
GPT-4.1 Nanoopenai/gpt-4.1-nano$0.10$0.40$0.0251M
GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch$0.05$0.20$0.0131M
o1-proopenai/o1-pro$150.00$600.00—200K
o3 Mini Highopenai/o3-mini-high$1.10$4.40$0.55200K
o3 Miniopenai/o3-mini$1.10$4.40$0.55200K
o3 Mini (batch)openai/o3-mini:batch$0.55$2.20$0.275200K
o1openai/o1$15.00$60.00$7.50200K
GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20$2.50$10.00$1.25128K
GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06$2.50$10.00$1.25128K
GPT-4o-miniopenai/gpt-4o-mini$0.15$0.60$0.075128K
GPT-4o-mini (2024-07-18)openai/gpt-4o-mini-2024-07-18$0.15$0.60$0.075128K
GPT-4o-mini (batch)openai/gpt-4o-mini:batch$0.075$0.30$0.037128K
GPT-4oopenai/gpt-4o$2.50$10.00$1.25128K
GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13$5.00$15.00—128K
GPT-4o (batch)openai/gpt-4o:batch$1.25$5.00$0.625128K
GPT-4 Turboopenai/gpt-4-turbo$10.00$30.00—128K
GPT-4 Turbo (batch)openai/gpt-4-turbo:batch$5.00$15.00—128K
GPT-3.5 Turbo (older v0613)openai/gpt-3.5-turbo-0613$1.00$2.00—4K
GPT-3.5 Turbo Instructopenai/gpt-3.5-turbo-instruct$1.50$2.00—4K
GPT-3.5 Turbo 16kopenai/gpt-3.5-turbo-16k$3.00$4.00—16K
GPT-3.5 Turboopenai/gpt-3.5-turbo$0.50$1.50—16K
GPT-3.5 Turbo (batch)openai/gpt-3.5-turbo:batch$0.25$0.75—16K
GPT-4openai/gpt-4$30.00$60.00—8K
Free Models Routeropenrouter/free$0$0—200K
Perceptron Mk1.5perceptron/perceptron-mk1.5$0.15$1.50—37K
Perceptron Mk1perceptron/perceptron-mk1$0.15$1.50—33K
Sonar Pro Searchperplexity/sonar-pro-search$3.00$15.00—200K
Sonar Reasoning Properplexity/sonar-reasoning-pro$2.00$8.00—128K
Sonar Properplexity/sonar-pro$3.00$15.00—200K
Sonar Deep Researchperplexity/sonar-deep-research$2.00$8.00—128K
Sonarperplexity/sonar$1.00$1.00—127K
Laguna S 2.1poolside/laguna-s-2.1$0.09$0.18$0.00901M
Laguna S 2.1 (free)poolside/laguna-s-2.1:free$0$0—262K
Laguna XS 2.1poolside/laguna-xs-2.1$0.06$0.12$0.03262K
Laguna XS 2.1 (free)poolside/laguna-xs-2.1:free$0$0—262K
Ternary Bonsai 2 27Bprism-ml/ternary-bonsai-2-27b$0.075$0.50$0.037262K
Qwen3.8 Max Primeqwen/qwen3.8-max-prime$4.00$12.00$0.501M
Qwen3.8 Omni Flashqwen/qwen3.8-omni-flash$0.15$0.47$0.0161M
Qwen3.8 Max (0902)qwen/qwen3.8-max-0902$2.00$6.00$0.251M
Qwen3.8 Flashqwen/qwen3.8-flash$0.15$0.47$0.0161M
Qwen3.8 27Bqwen/qwen3.8-27b$0.425$2.55$0.0851M
Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b$2.00$6.00$0.251M
Qwen3.7 Flashqwen/qwen3.7-flash$0.03$0.13$0.00601M
Qwen3.7 Plusqwen/qwen3.7-plus$0.32$1.28$0.0641M
Qwen3.7 Maxqwen/qwen3.7-max$1.48$4.42$0.2951M
Qwen3.5 Plus 2026-04-20qwen/qwen3.5-plus-20260420$0.30$1.80—1M
Qwen3.6 Flashqwen/qwen3.6-flash$0.188$1.13—1M
Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b$0.15$1.00$0.05262K
Qwen3.6 Max Previewqwen/qwen3.6-max-preview$1.03$6.16—262K
Qwen3.6 27Bqwen/qwen3.6-27b$0.30$2.00$0.03262K
Qwen3.6 Plusqwen/qwen3.6-plus$0.325$1.95—1M
Qwen3.5-9Bqwen/qwen3.5-9b$0.10$0.15—262K
Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3b$0.15$1.00$0.05262K
Qwen3.5-27Bqwen/qwen3.5-27b$0.26$2.60—262K
Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b$0.26$2.08—262K
Qwen3.5-Flashqwen/qwen3.5-flash-02-23$0.065$0.26—1M
Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15$0.26$1.56—1M
Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b$0.45$3.00$0.22262K
Qwen3 Max Thinkingqwen/qwen3-max-thinking$0.78$3.90—262K
Qwen3 Coder Nextqwen/qwen3-coder-next$0.12$0.80$0.07262K
Qwen3 VL 32B Instructqwen/qwen3-vl-32b-instruct$0.104$0.416—131K
Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking$0.18$2.10—131K
Qwen3 VL 8B Instructqwen/qwen3-vl-8b-instruct$0.117$0.455—262K
Qwen3 VL 30B A3B Thinkingqwen/qwen3-vl-30b-a3b-thinking$0.20$2.40—262K
Qwen3 VL 30B A3B Instructqwen/qwen3-vl-30b-a3b-instruct$0.15$0.60—262K
Qwen3 VL 235B A22B Thinkingqwen/qwen3-vl-235b-a22b-thinking$0.40$4.00—131K
Qwen3 VL 235B A22B Instructqwen/qwen3-vl-235b-a22b-instruct$0.21$1.90$0.10262K
Qwen3 Maxqwen/qwen3-max$0.78$3.90$0.156262K
Qwen3 Coder Plusqwen/qwen3-coder-plus$0.65$3.25$0.131M
Qwen3 Coder Flashqwen/qwen3-coder-flash$0.195$0.975$0.0391M
Qwen3 Next 80B A3B Thinkingqwen/qwen3-next-80b-a3b-thinking$0.15$1.20—262K
Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct$0.15$1.50—262K
Qwen Plus 0728qwen/qwen-plus-2025-07-28$0.26$0.78—1M
Qwen3 30B A3B Thinking 2507qwen/qwen3-30b-a3b-thinking-2507$0.20$2.40—82K
Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct$0.07$0.28—262K
Qwen3 30B A3B Instruct 2507qwen/qwen3-30b-a3b-instruct-2507$0.048$0.193—262K
Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507$0.23$2.30—131K
Qwen3 Coder 480B A35Bqwen/qwen3-coder$0.30$1.00$0.10262K
Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507$0.09$0.55—262K
Qwen3 30B A3Bqwen/qwen3-30b-a3b$0.12$0.50—131K
Qwen3 8Bqwen/qwen3-8b$0.117$0.455—131K
Qwen3 14Bqwen/qwen3-14b$0.12$0.24—131K
Qwen3 32Bqwen/qwen3-32b$0.08$0.28—131K
Qwen3 235B A22Bqwen/qwen3-235b-a22b$0.455$1.82—131K
Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct$0.80$1.00$0.40128K
Qwen-Plusqwen/qwen-plus$0.26$0.78$0.0521M
Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct$0.66$1.00—33K
Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct$0.10$0.20—33K
Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct$0.36$0.40—33K
Reka Edgerekaai/reka-edge$0.10$0.10—16K
Reka Flash 3rekaai/reka-flash-3$0.10$0.20—66K
Relace Searchrelace/relace-search$1.00$3.00—256K
Relace Apply 3relace/relace-apply-3$0.85$1.25—256K
Fugu Ultra v2sakana/fugu-ultra-v2$5.00$30.00$0.501M
Fugu Maxsakana/fugu-max$2.00$6.00$0.251M
Sakana Namazusakana/sakana-namazu$0.95$4.00$0.15262K
Fugu Ultrasakana/fugu-ultra$5.00$30.00$0.501M
Llama 3.3 Euryale 70Bsao10k/l3.3-euryale-70b$0.65$0.75—131K
Llama 3.1 Euryale 70B v2.2sao10k/l3.1-euryale-70b$0.85$0.85—131K
Llama 3 8B Lunarissao10k/l3-lunaris-8b$0.04$0.05—8K
Step 3.7 Flashstepfun/step-3.7-flash$0.20$1.15$0.04262K
Step 3.5 Flashstepfun/step-3.5-flash$0.10$0.30—262K
Hy4 previewtencent/hy4-preview$0.834$2.50$0.0421M
Hy-MT2-1.8Btencent/hy-mt2-1.8b$0.044$0.177—8K
Hy-MT2-30B-A3Btencent/hy-mt2-30b-a3b$0.074$0.295—8K
Hy-MT2-7Btencent/hy-mt2-7b$0.074$0.295—8K
Hy3tencent/hy3$0.132$0.528$0.033262K
Hy3 previewtencent/hy3-preview$0.18$0.60$0.06262K
Hunyuan A13B Instructtencent/hunyuan-a13b-instruct$0.14$0.57—131K
Cydonia 24B V4.1thedrummer/cydonia-24b-v4.1$0.30$0.50$0.15131K
Skyfall 36B V2thedrummer/skyfall-36b-v2$0.55$0.80$0.2533K
UnslopNemo 12Bthedrummer/unslopnemo-12b$0.40$0.40—1M
Inkling Smallthinkingmachines/inkling-small$0.45$1.20$0.10524K
Inkling Small (free)thinkingmachines/inkling-small:free$0$0—1M
Inklingthinkingmachines/inkling$1.00$4.05$0.17524K
Inkling (free)thinkingmachines/inkling:free$0$0—1M
Pareto 26.10 Previewunbiased/pareto-26.10-preview$0.80$3.20$0.031M
Paretounbiased/pareto$2.50$7.50$0.25262K
ReMM SLERP 13Bundi95/remm-slerp-l2-13b$0.35$0.65—6K
Solar Mini 4upstage/solar-mini4$0.05$0.20$0.0050524K
Solar Pro 4upstage/solar-pro4$0.09$0.36$0.018524K
Solar Pro 3upstage/solar-pro-3$0.15$0.60$0.015131K
Palmyra X5writer/palmyra-x5$0.60$6.00—1M
Grok 4.7x-ai/grok-4.7$2.00$6.00$0.50500K
Grok 4.6x-ai/grok-4.6$2.00$6.00$0.50500K
Grok 4.5x-ai/grok-4.5$2.00$6.00$0.30500K
Grok Build 0.1x-ai/grok-build-0.1$1.00$2.00$0.20256K
Grok 4.3x-ai/grok-4.3$1.25$2.50$0.201M
Grok 4.3 (batch)x-ai/grok-4.3:batch$1.00$2.00$0.161M
Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent$1.25$2.50$0.202M
Grok 4.20x-ai/grok-4.20$1.25$2.50$0.202M
MiMo-V2.6-Pro-UltraSpeedxiaomi/mimo-v2.6-pro-ultraspeed$4.35$8.70$0.0361M
MiMo-V2.6-Flashxiaomi/mimo-v2.6-flash$0.14$0.28$0.00281.1M
MiMo-V2.6-Proxiaomi/mimo-v2.6-pro$0.435$0.87$0.00361.1M
MiMo-V2.5-Proxiaomi/mimo-v2.5-pro$0.435$0.87$0.00361.1M
MiMo-V2.5xiaomi/mimo-v2.5$0.14$0.28$0.00281.1M
GLM 5.3 Primez-ai/glm-5.3-prime$2.80$8.80$0.561M
GLM 5.3 FlashXz-ai/glm-5.3-flashx$0.37$1.25$0.091M
GLM 5.3 Flashz-ai/glm-5.3-flash$0.15$0.50$0.031M
GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch$0.06$0.20$0.0121M
GLM 5.3z-ai/glm-5.3$0.049$3.39$0.0481M
GLM 5.3 (batch)z-ai/glm-5.3:batch$0.45$2.00$0.101M
GLM 5.2z-ai/glm-5.2$0.03$10.00$0.031M
GLM 5.1z-ai/glm-5.1$0.966$3.04$0.179205K
GLM 5V Turboz-ai/glm-5v-turbo$1.20$4.00$0.24203K
GLM 5 Turboz-ai/glm-5-turbo$1.20$4.00$0.24203K
GLM 5z-ai/glm-5$0.60$1.92$0.12205K
GLM 4.7 Flashz-ai/glm-4.7-flash$0.06$0.40—200K
GLM 4.7z-ai/glm-4.7$0.60$2.20$0.11205K
GLM 4.6Vz-ai/glm-4.6v$0.30$0.90$0.055131K
GLM 4.6z-ai/glm-4.6$0.43$1.75$0.08205K
GLM 4.5Vz-ai/glm-4.5v$0.60$1.80$0.1166K
GLM 4.5z-ai/glm-4.5$0.60$2.20$0.11131K
GLM 4.5 Airz-ai/glm-4.5-air$0.13$0.85$0.025131K

FAQ

How do you calculate the cost of an LLM API call?

Multiply input tokens by the model’s input price and output tokens by its output price, both per million tokens, and add them. Cached input tokens are billed at the lower cache-read price. Multiply by calls per day and by 30 for a monthly figure.

How many tokens is a word?

For English text, one token is about four characters, or roughly three quarters of a word. 1,000 words is about 1,300 tokens. Code and non-English text usually take more tokens per word.

Why are output tokens more expensive than input tokens?

Input tokens are processed in parallel in one pass, while output tokens are generated one at a time, so each output token costs the provider more compute. Most models price output at 3 to 8 times their input price.

Where do these prices come from?

From the public OpenRouter models API, refreshed every time trygin.ai is deployed. OpenRouter passes provider prices through without markup; it charges a fee when you buy credits. More in OpenRouter pricing explained.