Skip to main content
198 current AI models

AI API Cost Calculator

Calculate API costs for every model currently included in the GLLMPI. For time-based tariffs, the calculator uses the off-peak base price.

New: The model catalog is matched directly to the GLLMPI.

Filter & Search
Find the perfect model for your requirements
198 models found
Calculation
Enter your parameters to calculate costs

≈ 1000 Tokens

≈ 1000 Tokens

Total cost:$0.1400
Cost per call:$0.001400
Input cost:$0.0200
Output cost:$0.1200
Model Comparison
Cost comparison of all available AI models
ModelAPI ProviderInput $/1MOutput $/1MContextTotal Cost
inclusionAI$0.02$0.06262K
$0.0084
Granite 4.0 H Micro
IBM$0.02$0.11131K
$0.0129
Llama 3.1 8B
Meta$0.05$0.08131K
$0.0130
Qwen 3.7 Flash
Alibaba$0.03$0.111M
$0.0138
Granite 4.1 8B
IBM$0.05$0.10131K
$0.0150
Gemma 3 4B
Google$0.05$0.10128K
$0.0150
GPT OSS 20B
OpenAI$0.03$0.13131K
$0.0160
Amazon Nova Micro
Amazon$0.04$0.14128K
$0.0175
Laguna XS 2.1
Poolside$0.06$0.12262K
$0.0180
Command R7B
Cohere$0.04$0.15128K
$0.0187
GPT OSS 120B
OpenAI$0.03$0.17131K
$0.0200
Gemma 3 12B
Google$0.05$0.15128K
$0.0200
Ministral 3 3B
Mistral AI$0.10$0.10262K
$0.0200
Phi-4
Microsoft$0.07$0.1416K
$0.0210
Llama 3.2 1B
Meta$0.03$0.20131K
$0.0228
Qwen 3 30B A3B Instruct 2507
Alibaba$0.05$0.19262K
$0.0241
Nemotron 3 Nano
NVIDIA$0.05$0.201M
$0.0250
Qwen 3.5 9B
Alibaba$0.10$0.15262K
$0.0250
Laguna S 2.1
Poolside$0.09$0.181.05M
$0.0270
Nemotron 3.5 Lightning
NVIDIA$0.08$0.201M
$0.0280
Nemotron Nano 2 9B
NVIDIA$0.06$0.23128K
$0.0290
Amazon Nova Lite
Amazon$0.06$0.24300K
$0.0300
Solar Mini
Upstage$0.15$0.1532K
$0.0300
Ministral 3 8B
Mistral AI$0.15$0.15262K
$0.0300
Qwen 2.5 7B
Alibaba$0.10$0.20131K
$0.0300
Qwen 3 Coder 30B A3B
Alibaba$0.07$0.28262K
$0.0350
Qwen 3 32B
Alibaba$0.08$0.2832K
$0.0360
Qwen 3 14B
Alibaba$0.12$0.2432K
$0.0360
Seed 1.6 Flash
ByteDance$0.07$0.30256K
$0.0375
Llama 3.2 3B
Meta$0.05$0.33131K
$0.0380
Llama 4 Scout
Meta$0.10$0.3010.49M
$0.0400
Ministral 3 14B
Mistral AI$0.20$0.20262K
$0.0400
Voxtral Small 24B
Mistral AI$0.10$0.3032K
$0.0400
Step 3.5 Flash
StepFun$0.10$0.30256K
$0.0400
Gemma 4 26B A4B
Google$0.07$0.34262K
$0.0410
Llama 3.3 70B
Meta$0.10$0.32131K
$0.0420
MiMo-V2.5
Xiaomi$0.14$0.281M
$0.0420
Gemma 4 31B
Google$0.10$0.34262K
$0.0440
GLM-4.7-Flash
Z.ai$0.06$0.40200K
$0.0460
Nemotron 3 Super
NVIDIA$0.09$0.401M
$0.0485
Qwen 3.8 Flash
Alibaba$0.11$0.381M
$0.0495
Gemini 2.5 Flash-Lite
Google$0.10$0.401.05M
$0.0500
Seed 2.0 Mini
ByteDance$0.10$0.40256K
$0.0500
Qwen 3 VL 32B
Alibaba$0.10$0.42262K
$0.0520
Gemma 3 27B
Google$0.08$0.45128K
$0.0530
Qwen 3 8B
Alibaba$0.12$0.4632K
$0.0572
Qwen 3 VL 8B
Alibaba$0.12$0.46262K
$0.0572
GPT-6 Luna
OpenAI$0.10$0.501.05M
$0.0600
DeepSeek-V3.2
DeepSeek$0.26$0.38128K
$0.0640
Qwen 3 235B A22B Instruct 2507
Alibaba$0.09$0.55262K
$0.0640
Qwen 3 30B A3B
Alibaba$0.13$0.5232K
$0.0650
Qwen 3 VL 30B A3B
Alibaba$0.13$0.52262K
$0.0650
GLM-5.3-Flash
Z.ai$0.15$0.501M
$0.0650
Hy3
Tencent$0.13$0.53262K
$0.0660
DeepSeek-V4-Flash
DeepSeek$0.04$0.641M
$0.0680
DeepSeek-V3.2 Exp
DeepSeek$0.27$0.41128K
$0.0680
Hunyuan A13B Instruct
Tencent$0.14$0.57262K
$0.0710
DeepSeek-V4.1-Flash(Off-peak base)
DeepSeek$0.15$0.601M
$0.0750
GPT-4o mini
OpenAI$0.15$0.60128K
$0.0750
Mistral Small 4
Mistral AI$0.15$0.60256K
$0.0750
Solar Pro 2
Upstage$0.15$0.6065K
$0.0750
Solar Pro 3
Upstage$0.15$0.60131K
$0.0750
Qwen 2.5 72B
Alibaba$0.36$0.40131K
$0.0760
Nemotron Nano 2 VL 12B
NVIDIA$0.20$0.60128K
$0.0800
Llama 3.1 70B
Meta$0.40$0.40131K
$0.0800
Qwen 3 Coder Next
Alibaba$0.12$0.80262K
$0.0920
GLM-4.5 Air
Z.ai$0.13$0.85131K
$0.0980
Llama 4 Maverick
Meta$0.20$0.801.05M
$0.1000
Trinity Large Thinking
Arcee AI$0.22$0.85262K
$0.1070
Qwen 3.6 35B A3B
Alibaba$0.14$1.00262K
$0.1140
Qwen 3 Next 80B A3B Instruct
Alibaba$0.10$1.10262K
$0.1200
GLM-4.6V
Z.ai$0.30$0.90131K
$0.1200
DeepSeek-V3 0324
DeepSeek$0.25$1.00163K
$0.1250
DeepSeek-V3.1 Terminus
DeepSeek$0.27$1.00163K
$0.1270
MiniMax M2
MiniMax$0.26$1.02204K
$0.1275
DeepSeek-V3
DeepSeek$0.26$1.03128K
$0.1286
Qwen 3 Coder 480B A35B
Alibaba$0.30$1.00262K
$0.1300
MiniMax-01
MiniMax$0.20$1.101M
$0.1300
Gemma 2 27B
Google$0.65$0.658K
$0.1300
MiMo-V2.5-Pro
Xiaomi$0.43$0.871M
$0.1305
Step 3.7 Flash
StepFun$0.20$1.15256K
$0.1350
MiniMax M2.5
MiniMax$0.27$1.08204K
$0.1350
Qwen 3 Next 80B A3B Thinking
Alibaba$0.15$1.20262K
$0.1350
Qwen 3.7 Plus
Alibaba$0.28$1.101M
$0.1377
GPT-5.6 Luna
OpenAI$0.20$1.201.05M
$0.1400
Muse Glimmer 30B
Meta$0.30$1.10131K
$0.1400
MiniMax M2.7
MiniMax$0.30$1.20204K
$0.1500
Qwen 3.5 35B A3B
Alibaba$0.25$1.25262K
$0.1500
MiniMax M2.1
MiniMax$0.30$1.20204K
$0.1500
MiniMax M3
MiniMax$0.30$1.201M
$0.1500
Inkling Small
Thinking Machines Lab$0.30$1.201M
$0.1500
Solar Pro 4
Upstage$0.30$1.20512K
$0.1500
Qwen 2.5 Coder 32B
Alibaba$0.66$1.00131K
$0.1660
ERNIE 4.5 VL 424B A47B
Baidu$0.42$1.25131K
$0.1670
Qwen 3.5 27B
Alibaba$0.20$1.56262K
$0.1755
Qwen 2.5 VL 72B
Alibaba$0.80$1.00128K
$0.1800
GPT-4.1 mini
OpenAI$0.40$1.601.05M
$0.2000
Sonar
Perplexity$1.00$1.00128K
$0.2000
Mistral Large 3
Mistral AI$0.50$1.50262K
$0.2000
Qwen 3 VL 235B A22B
Alibaba$0.21$1.90262K
$0.2110
GLM-4.7
Z.ai$0.40$1.75204K
$0.2150
DeepSeek-V3.1
DeepSeek$0.55$1.65128K
$0.2200
Seed 1.6
ByteDance$0.25$2.00256K
$0.2250
Seed 2.0 Lite
ByteDance$0.25$2.00256K
$0.2250
Seed 1.8
ByteDance$0.25$2.00256K
$0.2250
Qwen 3 235B A22B
Alibaba$0.46$1.8232K
$0.2275
Qwen 3 VL 8B Thinking
Alibaba$0.18$2.10262K
$0.2280
Qwen 3.5 122B A10B
Alibaba$0.26$2.08262K
$0.2340
GLM-4.5V
Z.ai$0.60$1.8065K
$0.2400
GLM-4.6
Z.ai$0.50$2.00204K
$0.2500
Qwen 3 235B A22B Thinking 2507
Alibaba$0.23$2.30262K
$0.2530
Qwen 3 30B A3B Thinking 2507
Alibaba$0.20$2.40262K
$0.2600
Qwen 3 VL 30B A3B Thinking
Alibaba$0.20$2.40262K
$0.2600
DeepSeek-V4-Pro(Off-peak base)
DeepSeek$0.66$1.981M
$0.2640
DeepSeek-R1 0528
DeepSeek$0.50$2.15163K
$0.2650
Kimi K2.5
Moonshot AI$0.45$2.25262K
$0.2700
Qwen 3.5 397B A17B
Alibaba$0.39$2.34262K
$0.2730
MiniMax M1
MiniMax$0.55$2.201M
$0.2750
Gemini 3.5 Flash-Lite
Google$0.30$2.501.05M
$0.2800
Amazon Nova 2 Lite
Amazon$0.30$2.501M
$0.2800
Ring 2.6 1T
inclusionAI$0.30$2.50262K
$0.2800
GLM-4.5
Z.ai$0.60$2.20131K
$0.2800
Gemini 2.5 Flash
Google$0.30$2.501.05M
$0.2800
Kimi K2 0711
Moonshot AI$0.57$2.30131K
$0.2870
Seed 2.1 Turbo
ByteDance$0.50$2.50256K
$0.3000
Grok Build 0.1
xAI$1.00$2.00256K
$0.3000
Kimi K2 Thinking
Moonshot AI$0.60$2.50262K
$0.3100
Kimi K2 0905
Moonshot AI$0.60$2.50262K
$0.3100
DeepSeek-R1
DeepSeek$0.70$2.50128K
$0.3200
Gemini 3 Flash Preview
Google$0.50$3.001.05M
$0.3500
Seed 2.0 Code
ByteDance$0.50$3.00256K
$0.3500
Seed 2.0 Pro
ByteDance$0.50$3.00256K
$0.3500
Qwen 3.8 27B
Alibaba$0.45$3.20262K
$0.3650
KAT-Coder-Pro V2.5
Kuaishou$0.74$2.96256K
$0.3700
LongCat 2.0
Meituan$0.75$2.951M
$0.3700
Grok 4.3
xAI$1.25$2.501M
$0.3750
Grok 4.20 Reasoning
xAI$1.25$2.501M
$0.3750
Grok 4.20 Multi-Agent
xAI$1.25$2.501M
$0.3750
MiMo-V2.5-Pro-UltraSpeed
Xiaomi$1.30$2.611M
$0.3915
Amazon Nova Pro
Amazon$0.80$3.20300K
$0.4000
Hermes 4 405B
Nous Research$1.00$3.00131K
$0.4000
Nemotron 3 Ultra
NVIDIA$0.60$3.601M
$0.4200
Qwen 3.6 27B
Alibaba$0.60$3.60262K
$0.4200
GLM-5
Z.ai$1.00$3.20200K
$0.4200
Qwen 3 VL 235B A22B Thinking
Alibaba$0.40$4.00262K
$0.4400
Gemini 3.8 Flash
Google$0.75$3.751.05M
$0.4500
Gemini 3.7 Flash
Google$0.75$3.751.05M
$0.4500
Gemini 3.6 Flash
Google$0.75$3.751.05M
$0.4500
Kimi K2.6
Moonshot AI$0.95$4.00262K
$0.4950
Kimi K2.7 Code
Moonshot AI$0.95$4.00262K
$0.4950
Inkling
Thinking Machines Lab$1.00$4.051M
$0.5050
GPT-5.4 mini
OpenAI$0.75$4.50400K
$0.5250
Muse Spark 1.3
Meta$1.25$4.251.05M
$0.5500
GLM-5.2
Z.ai$1.40$4.401M
$0.5800
GLM-5.3
Z.ai$1.40$4.401M
$0.5800
GLM-5.1
Z.ai$1.40$4.40200K
$0.5800
Claude Haiku 4.5
Anthropic$1.00$5.00200K
$0.6000
Qwen 3.7 Max
Alibaba$1.65$4.951M
$0.6601
Qwen 3.8 Max 0902
Alibaba$1.65$4.951M
$0.6601
Qwen 3.8 2.4T A95B
Alibaba$2.00$6.00262K
$0.8000
Mixtral 8x22B
Mistral AI$2.00$6.0065K
$0.8000
Grok 4.7
xAI$2.00$6.00500K
$0.8000
Grok 4.6
xAI$2.00$6.00500K
$0.8000
Grok 4.5
xAI$2.00$6.00500K
$0.8000
Mistral Medium 3.5
Mistral AI$1.50$7.50262K
$0.9000
GPT-4.1
OpenAI$2.00$8.001.05M
$1.0000
Sonar Reasoning Pro
Perplexity$2.00$8.00128K
$1.0000
Sonar Deep Research
Perplexity$2.00$8.00128K
$1.0000
Gemini 3.5 Flash
Google$1.50$9.001.05M
$1.0500
Gemini 2.5 Pro
Google$1.25$10.001.05M
$1.1250
Claude Sonnet 5.5
Anthropic$2.00$10.001M
$1.2000
Claude Sonnet 5
Anthropic$2.00$10.001M
$1.2000
GPT-6 Sol
OpenAI$2.00$10.001.05M
$1.2000
GPT-4o
OpenAI$2.50$10.00128K
$1.2500
Command A
Cohere$2.50$10.00256K
$1.2500
GPT-5.6 Terra
OpenAI$2.00$12.001.05M
$1.4000
Gemini 3.1 Pro Preview
Google$2.00$12.001.05M
$1.4000
Amazon Nova Premier
Amazon$2.50$12.501M
$1.5000
GPT-5.4
OpenAI$2.50$15.001.05M
$1.7500
Claude Sonnet 4.6
Anthropic$3.00$15.001M
$1.8000
Claude Sonnet 4.5
Anthropic$3.00$15.00200K
$1.8000
Sonar Pro
Perplexity$3.00$15.00200K
$1.8000
Kimi K3
Moonshot AI$3.00$15.001.05M
$1.8000
Claude Opus 5.5
Anthropic$4.00$20.001M
$2.4000
GPT-5.6 Sol
OpenAI$4.00$20.001.05M
$2.4000
Claude Opus 5
Anthropic$5.00$25.001M
$3.0000
Claude Opus 4.7
Anthropic$5.00$25.001M
$3.0000
Claude Opus 4.6
Anthropic$5.00$25.001M
$3.0000
Claude Opus 4.5
Anthropic$5.00$25.00200K
$3.0000
ChatGPT chat-latest
OpenAI$5.00$30.00400K
$3.5000
GPT-5.5
OpenAI$5.00$30.001.05M
$3.5000
Claude Fable 5.1
Anthropic$10.00$50.001M
$6.0000
Claude Mythos 5.1
Anthropic$10.00$50.001M
$6.0000
Claude Fable 5
Anthropic$10.00$50.001M
$6.0000
Claude Mythos 5
Anthropic$10.00$50.001M
$6.0000
GPT-6 Astra
OpenAI$10.00$50.001.05M
$6.0000
GPT-5.5 Pro
OpenAI$30.00$180.001.05M
$21.0000
GPT-5.4 Pro
OpenAI$30.00$180.001.05M
$21.0000

Note on Pricing

The displayed prices use the same continuously verified data as the GLLMPI. For time-based tariffs, the shown input and output prices are the off-peak base rates. The calculator does not switch by the clock, so higher rates may apply during peak hours. Caching, batch processing, volume discounts, context tiers, and additional API features can change your actual costs.

Frequently Asked Questions

Important information about using AI APIs

A token is a unit of text processed by the API. On average: 1 token ≈ 4 characters in English, 1 token ≈ ¾ words, 100 tokens ≈ 75 words. The exact token count can vary depending on language and text type.

The right model depends on your use case. Start with a cost comparison using your real token volume, then filter by provider, context window, and total cost. Lower-cost models suit high-volume work. For complex tasks or agent runs, a stronger model may be the better choice.

No API wins every task. First compare costs for your own prompts and typical outputs. Then account for context windows, latency, data handling, and the capabilities your product actually needs.

Shorter prompts, sensible output limits, and a model that fits each step can reduce costs substantially. Also check caching for repeated input plus your provider's batch and volume discounts.

Input tokens are all tokens you send to the API (your prompt, system message, etc.). Output tokens are the tokens in the API's response. Output tokens are typically more expensive because they need to be generated by the API, while input tokens are only processed.

OpenAI, Anthropic, and Google each offer a 50% discount for batch workloads that do not need immediate responses. Anthropic and OpenAI also support caching for repeated input. Google provides a free developer tier for selected Gemini models. Contract and volume discounts depend on the provider, workload, and negotiated agreement.

The calculator uses base prices from the current GLLMPI catalog. For time-based tariffs, that means the off-peak base price, not necessarily the tariff at the current time. The result remains an estimate because tokenization, context length, caching, batch processing, and contract terms can change your actual costs. Your provider's price is always the binding figure for billing.
Token Conversion & Examples
Understand how tokens are counted

Average Conversions

1 Token≈ 4 characters (English)
1 Token≈ 2-3 characters (German)
1 Word≈ 1.3 tokens
1 Sentence (15 words)≈ 20 tokens
1 Paragraph (100 words)≈ 130 tokens

Typical Use Cases

Simple Q&A
Input: ~50 tokens, Output: ~100 tokens

GPT-5.6 Luna | Gemini 3.5 Flash-Lite | Claude Haiku 4.5

Text Summary
Input: ~1000 tokens, Output: ~200 tokens

GPT-5.6 Luna | Gemini 3.5 Flash-Lite | DeepSeek V4 Flash

Code Generation
Input: ~200 tokens, Output: ~500 tokens

GPT-5.6 Terra | Claude Sonnet 5.5

Complex Reasoning
Input: ~500 tokens, Output: ~1000 tokens

GPT-5.6 Sol | Claude Opus 5.5

Changelog

The latest updates and improvements to our API Cost Calculator
v3.9.0September 30, 2026

Added Claude Sonnet 5.5

  • Added Claude Sonnet 5.5 at $2 input and $10 output per 1M tokens
  • Updated the model recommendations for code generation and complex reasoning to Claude Sonnet 5.5 and Claude Opus 5.5
v3.8.0September 3, 2026

Added Muse Spark 1.3 and Qwen models

  • Added Muse Spark 1.3 with the current Meta endpoint rate offered through OpenRouter
  • Updated Qwen 3.8 Max to the 0902 snapshot and current global Alibaba rate
  • Added Qwen 3.8 Flash at $0.113 input and $0.382 output per 1M tokens
  • Avoided counting Qwen 3.8 Flash Next twice because QwenCloud serves the open weights under the same API name
  • Omitted Gemini 3.8 Flash Cyber because Google does not publish a public API price
v3.6.0September 2, 2026

Added Gemini 3.8 Flash

  • Added Gemini 3.8 Flash with its official standard price
  • Added the current input and output rates through the end of 2026
v3.5.0August 22, 2026

Synchronized the catalog with the GLLMPI

  • Shows only models currently included in the GLLMPI
  • Builds the provider filter automatically from the catalog
  • Replaces external favicons with local provider icons
v3.4.0August 16, 2026

Added Gemini 3.7 Flash and Grok 4.6

  • Added Gemini 3.7 Flash with its current introductory price
  • Added Grok 4.6 with long-context pricing tiers
  • Expanded the calculator to 95 models
v3.3.0August 9, 2026

Added current models and pricing

  • Added Claude Opus 5, Gemini 3.6 Flash, and Gemini 3.5 Flash-Lite
  • Rechecked the official OpenAI, Anthropic, and Google pricing sources
  • Expanded the calculator to 93 models
v3.2.0July 13, 2026

Empty results message

  • Shows a no-models-found message with a reset action when filters exclude all models
  • Screen reader label for the clear-search button
v3.1.0March 13, 2026

UI improvements and centralized pricing

  • Migrated all model prices to centralized database
  • Added provider icons next to each model name
  • Removed badges and model descriptions (cleaner table)
  • Added Gemini 3.1 Flash-Lite
  • Total of 65+ AI models available
v3.0.0February 20, 2026

New Models: Claude 4.6, Gemini 3 Series

  • Added Claude Opus 4.6 with 1M context and agent teams
  • Added Claude Sonnet 4.6 with Opus-level performance
  • Added Claude Opus 4.5 with extended thinking
  • Added Gemini 3.1 Pro with GPQA Diamond record
  • Added Gemini 3 Pro and Gemini 3 Flash
  • Total of 65+ AI models available
v2.0.0September 15, 2025

Major Update with New Models

  • Added GPT-5 series (Nano, Mini, Standard, Chat)
  • Claude 4.1 Opus with 74.5% SWE-bench score
  • Added Claude 4.5 Sonnet
  • xAI Grok models (Grok 3 & 4 series)
  • Meta Llama 4 models (Scout & Maverick)
  • Added Mistral Large 2 and Small
  • Updated Gemini 2.5 Pro pricing
  • Total of 50+ AI models available
v1.2.1July 23, 2025

Changelog Added

  • Added changelog section to API cost calculator
v1.2.0July 22, 2025

Major Update

  • Added all Claude 4 and Gemini 2.5 models
  • Implemented comprehensive filter and sort functions
  • Extended comparison options for 36 AI models
v1.1.0July 22, 2025

New Models

  • Added GPT-4.1 series with 1M token context
  • Updated O3 model with 80% price reduction
  • Integrated new audio and realtime model variants
v1.0.0July 22, 2025

Initial Release

  • Initial version with 21 OpenAI models
  • Token conversion for words and characters
  • Interactive cost calculation and comparison table