Which model here is cheapest, and what does each one cost?

32 models are published here with live token prices, across 3 routing prefixes, and every rate is the provider's own with no margin added. The cheapest input rate is Gemini 2.5 Flash Lite at $0.10 per million input tokens and $0.40 per million output; the most expensive is Claude Fable 5 at $10.00 input. The largest published context window is 1,048,576 tokens on Gemini 2.5 Flash Lite. Prices are per million tokens in USD and are read from the upstream catalog twice an hour, so the table below is the authoritative figure rather than anything quoted elsewhere. These figures were last read from upstream on 2026-08-18. Credit is prepaid: what you top up becomes the spending ceiling on your API key, and one balance pays for any mixture of these models through a single OpenAI-compatible endpoint.

ModelModel idRouteContextInput / 1MOutput / 1M
Gemini 2.5 Flash Liteantigravity/gemini-2.5-flash-liteantigravity1,048,576$0.10$0.40
GPT-OSS 120B (Medium)antigravity/gpt-oss-120b-mediumantigravity131,072$0.10$0.40
GPT 5.6 Luna · 5 effort tierscx/gpt-5.6-lunacx272,000$0.20$1.20
Gemini 2.5 Flash Thinkingantigravity/gemini-2.5-flash-thinkingantigravity1,048,576$0.30$2.50
Gemini 2.5 Flashantigravity/gemini-2.5-flashantigravity1,048,576$0.30$2.50
Gemini 3.1 Flash Liteantigravity/gemini-3.1-flash-liteantigravity1,048,576$0.45$2.45
Claude 4.5 Haikucc/claude-haiku-4-5-20251001cc200,000$1.00$5.00
GPT 5.6 Luna (xHigh)cx/gpt-5.6-luna-xhighcx272,000$1.00$6.00
Gemini 3.5 Flash (High)antigravity/gemini-3-flash-agentantigravity1,048,576$1.50$9.00
Gemini 3.5 Flash (Medium)antigravity/gemini-3.5-flash-lowantigravity1,048,576$1.50$9.00
Gemini 3.5 Flash (Low)antigravity/gemini-3.5-flash-extra-lowantigravity1,048,576$1.50$9.00
Claude Sonnet 5cc/claude-sonnet-5cc1,000,000$2.00$10.00
Gemini 3.6 Flash (High) · 3 effort tiersantigravity/gemini-3.6-flash-highantigravity1,048,576$2.00$7.00
Gemini 3.1 Pro (High)antigravity/gemini-pro-agentantigravity1,048,576$2.00$12.00
Gemini 3.1 Pro (Low)antigravity/gemini-3.1-pro-lowantigravity1,048,576$2.00$12.00
GPT 5.6 Terra · 6 effort tierscx/gpt-5.6-terracx272,000$2.00$12.00
GPT 5.6 Terra (xHigh)cx/gpt-5.6-terra-xhighcx272,000$2.50$15.00
Claude 4.6 Sonnetcc/claude-sonnet-4-6cc1,000,000$3.00$15.00
Claude 4.5 Sonnetcc/claude-sonnet-4-5-20250929cc200,000$3.00$15.00
Claude Sonnet 4.6 (Thinking)antigravity/claude-sonnet-4-6antigravity1,048,576$3.00$15.00
Claude Opus 5cc/claude-opus-5cc1,000,000$5.00$25.00
Claude Opus 4.8cc/claude-opus-4-8cc1,000,000$5.00$25.00
Claude Opus 4.7cc/claude-opus-4-7cc1,000,000$5.00$25.00
Claude Opus 4.6cc/claude-opus-4-6cc1,000,000$5.00$25.00
Claude Opus 4.5cc/claude-opus-4-5-20251101cc200,000$5.00$25.00
Claude Opus 4.6 (Thinking)antigravity/claude-opus-4-6-thinkingantigravity1,048,576$5.00$25.00
GPT 5.6 Sol · 6 effort tierscx/gpt-5.6-solcxNot published$5.00$30.00
GPT 5.6 Sol (xHigh)cx/gpt-5.6-sol-xhighcxNot published$5.00$30.00
GPT 5.5 · 4 effort tierscx/gpt-5.5cx272,000$5.00$30.00
GPT 5.5 (xHigh)cx/gpt-5.5-xhighcx272,000$5.00$30.00
GPT 5.3 Codex Sparkcx/gpt-5.3-codex-sparkcxNot published$5.00$20.00
Claude Fable 5cc/claude-fable-5cc1,000,000$10.00$50.00
DoGPT

AI Model Pricing and Token Rates Catalog Guide

DoGPT passes model pricing straight through with zero markup added on top. Your credit is prepaid: the balance you top up becomes the exact spending ceiling on your API key. No subscriptions. No seat fees. No minimum spend requirements. Direct passthrough pricing.

We publish every rate openly without requiring an account.

We list exact input and output rates per million tokens across offered routes — including models from OpenAI, Anthropic, and Google under cc/, antigravity/, agy/, and cx/ prefixes. You can inspect context windows up to 1,048,576 tokens, maximum completion limits, and capability flags for vision, tool calling, reasoning, and thinking. Because we operate on a direct passthrough model, you pay only for the exact tokens your code consumes during inference.

Model IDs match upstream API specifications exactly. Point your OpenAI or Anthropic SDK code to our endpoint, drop in your key, and start running completions. The live table below refreshes twice every hour directly from upstream provider status.

Our cron reconciles token usage against your prepaid balance on regular ticks. You can monitor per-model cost breakdowns, request counts, and spend trends in your dashboard.

Inference calls connect directly to upstream provider endpoints, so this website never sits in your request path. Upstream enforces your spend limit, and we absorb any slight overshoot while an active request finishes so your balance never goes negative.

You can filter the catalog by connection type, search for specific model capabilities, or sort by token pricing directly in your browser. All published rates, context metrics, and provider availability flags refresh continuously.