Published Developer Rate Cards

AI Token Price Tracker & Cost Calculator

Compare stored public API rates and estimate simple usage where complete fixed USD input and output token prices are available.

Estimate your monthly token bill

Share or bookmark your inputs. Estimates use current catalog prices when opened, so future totals may change. Anyone with the link can see these volumes.

Illustrative assumptions, not measured averages. Edit every value to match your workload. Selecting an example resets its values.

Count instructions, conversation history, retrieved text and tool results each time they are sent. Include billed reasoning tokens in output where applicable.

4,500,000 input + 900,000 output tokens / month

Monthly tasks × calls per task × tokens per call

Standard USD token rates only. Cache discounts, batch/off-peak discounts, long-context tiers, tools, storage and taxes are excluded. Limits are screened using your averages; individual larger calls may need a different model or rate.

Lowest estimated token cost in this catalog

Gemini 2.5 Flash-Lite

$0.81 / month

Input
$0.45
Output
$0.36
Per task
$0.00081

A price comparison, not a quality recommendation. The catalog does not include every provider or model.

Read the price-versus-capability guide

Complete Rate Card Matrix

Model & FamilyCompanyInput / 1MOutput / 1MContext WindowEst. USD CostDetails
GPT-5.6 TerraReasoning & MultimodalAbove 272,000 tokens · Above 272K input tokens, the full request uses 2× input and 1.5× output pricing. · Cache writes are billed at 1.25× the uncached input rate.OpenAI$2.00$12.001,050,000 tokens$19.80Specs →
Muse Spark 1.2Agentic MultimodalMetaRate not verified in this recordRate not verified in this recordNot publishedComplete verified USD rates are required.Specs →
Gemini 2.5 Flash-LiteAI ModelStandard paid text, image, and video input is $0.10 per 1M tokens; audio input is $0.30 per 1M tokens. · Standard output is $0.40 per 1M tokens, including thinking tokens.Google$0.1000$0.40001,048,576 tokens$0.81Specs →
Gemini 3.1 Pro PreviewReasoning & MultimodalAbove 200,000 tokens · Standard pricing uses the higher tier above 200K prompt tokens; cached input is $0.40/MTok in that tier.Google$2.00$12.001,048,576 tokens$19.80Specs →
GPT-5.6 SolReasoning & MultimodalAbove 272,000 tokens · GPT-5.6 Sol’s $4 input / $20 output promotional pricing is available at least through November 21, 2026. · Above 272K input tokens, the full request uses 2× input and 1.5× output pricing. · Cache writes are billed at 1.25× the uncached input rate.OpenAI$4.00$20.001,050,000 tokens$36.00Specs →
Gemini 3.8 Flash CyberCybersecurity AIGoogleRate not verified in this recordRate not verified in this recordNot publishedComplete verified USD rates are required.Specs →
Claude Fable 5Reasoning & MultimodalAnthropic$10.00$50.001M tokens$90.00Specs →
GPT-5.6 LunaEfficient MultimodalAbove 272,000 tokens · Above 272K input tokens, the full request uses 2× input and 1.5× output pricing. · Cache writes are billed at 1.25× the uncached input rate.OpenAI$0.2000$1.201,050,000 tokens$1.98Specs →
GPT-6 AstraAI ModelAbove 272,000 tokens · Standard text-token rates. Above 272K input tokens, the entire request uses 2× input/cache and 1.5× output rates. · Cache writes cost 1.25× the uncached input rate. Batch/Flex cost 50% of Standard; Fast mode costs 2× applicable rates. Tool fees are additional.OpenAI$10.00$50.001,050,000 tokens$90.00Specs →
Gemini 3.7 FlashFast MultimodalPublished through Dec 31, 2026 · These are standard paid-tier introductory rates through December 31, 2026.Google$0.7500$3.751,048,576 tokens$6.75Specs →
Grok 4.6Reasoning & MultimodalAt or above 200,000 tokens · At or above 200K prompt tokens, the entire request uses $4/MTok input, $1/MTok cached input, and $12/MTok output rates. · The US regional endpoint charges a 10% premium on token usage, including long-context rates.SpaceXAI$2.00$6.00500K tokens$14.40Specs →
GPT-6 LunaReasoning & MultimodalOpenAI$0.1000$0.50001,050,000 tokens$0.90Specs →
DeepSeek V4 ProReasoning & CodingDisplayed rates use peak-hour pricing: input cache miss $1.32, cache hit $0.044, output $3.96 per 1M tokens. · Off-peak rates are $0.66 input cache miss, $0.022 cache hit and $1.98 output per 1M tokens. · Peak hours: Monday–Friday 01:00–04:00 and 06:00–10:00 UTC. All other hours are off-peak. The calculator uses peak rates and does not schedule traffic.DeepSeek$1.32$3.961M tokens$9.50Specs →
Claude Mythos 5.1AI ModelAnthropic$10.00$50.001M tokens$90.00Specs →
GPT-6 SolReasoning & MultimodalOpenAI$2.00$10.001,050,000 tokens$18.00Specs →
Kimi K3Reasoning & MultimodalRegion: China platform (platform.kimi.com) · Official China-platform rates are CNY-denominated. China and international Kimi API accounts, balances, and keys are separate, so these rates should not be treated as universal international pricing.Moonshot AICN¥20.00CN¥100.001,048,576 tokensComplete verified USD rates are required.Specs →
Mistral Small 4AI ModelMistral AI$0.1500$0.6000256K tokens$1.21Specs →
Mistral Medium 3.5Multimodal & CodingMistral AI$1.50$7.50256K tokens$13.50Specs →
GLM-5.3Reasoning & CodingZ.aiRate not verified in this recordRate not verified in this record1M tokensComplete verified USD rates are required.Specs →
Gemini 3.8 FlashFast MultimodalPublished through Dec 31, 2026 · Introductory paid-tier rates apply through December 31, 2026. Starting January 1, 2027, standard input/output pricing rises to $1.50/$7.50 per 1M tokens.Google$0.7500$3.751,048,576 tokens$6.75Specs →
Qwen3.8-MaxReasoning & MultimodalRegion: US (Virginia) · Global scope · Cached input uses the documented implicit-cache rate for the US (Virginia), Global-scope deployment.Alibaba Cloud$1.65$4.951,000,000 tokens$11.88Specs →
Claude Fable 5.1AI ModelAnthropic$10.00$50.001M tokens$90.00Specs →
Claude Sonnet 5.5Reasoning & AgenticAnthropic reports standard API pricing of $2 per million input tokens, $10 per million output tokens, $0.20 per million cache reads, and $2.50 per million cache writes.Anthropic$2.00$10.001,000,000 tokens$18.00Specs →
Claude Sonnet 5Reasoning & CodingAnthropic$2.00$10.001M tokens$18.00Specs →
Claude Opus 5Reasoning & CodingAnthropic$5.00$25.001M tokens$45.00Specs →