TokenGate pricing
AI API pricing comparison
Compare TokenGate prices for current OpenAI, Anthropic, Google, DeepSeek and xAI models. Input, cached-input and output prices are shown per 1 million tokens in USDT unless a model uses tiered pricing.
- Current models
- 20
- Providers
- 5
- Token pricing
- Per 1M tokens
- Billing
- USDT
LLM API price comparison
Compare current TokenGate model prices side by side. Select a provider, search for a model or sort by token cost and context window.
All current prices are TokenGate customer prices in USDT.
| Model | Provider | Input / 1M | Cached input / 1M | Output / 1M | Context | Details |
|---|---|---|---|---|---|---|
| GPT-6 Astra | OpenAI | $10.00 | $1.00 | $50.00 | 1.05M | Pricing details → |
| GPT-5.6 Luna | OpenAI | $0.20 | $0.020 | $1.20 | 1.05M | Pricing details → |
| GPT-5.6 Sol | OpenAI | $4.00 | $0.40 | $20.00 | 1.05M | Pricing details → |
| GPT-5.6 Terra | OpenAI | $2.00 | $0.20 | $12.00 | 1.05M | Pricing details → |
| Claude Fable 5.1 | Anthropic | $10.00 | $0.25 | $50.00 | — | Pricing details → |
| Claude Sonnet 5 | Anthropic | $2.00 | — | $10.00 | — | Pricing details → |
| Claude Opus 5 | Anthropic | $5.00 | — | $25.00 | — | Pricing details → |
| Grok 4.6Tiered | xAI | $2.00–$4.00 | $0.50–$1.00 | $6.00–$12.00 | 500K | Pricing details → |
| DeepSeek V4.1 Flash | DeepSeek | Variable | — | Variable | — | View pricing tiers → |
| Gemini 3.8 FlashScheduled rate change | $0.75 | $0.075 | $3.75 | 1.05M | Pricing details → | |
| Claude Opus 4.8 | Anthropic | $5.00 | — | $25.00 | — | Pricing details → |
| Claude Sonnet 4.6 | Anthropic | $3.00 | — | $15.00 | 1M | Pricing details → |
| Claude Haiku 4.5 | Anthropic | $1.00 | — | $5.00 | — | Pricing details → |
| Grok 4.5Tiered | xAI | $2.00–$4.00 | $0.30–$0.60 | $6.00–$12.00 | 500K | Pricing details → |
| Grok 4.3Tiered | xAI | $1.25–$2.50 | $0.20–$0.40 | $2.50–$5.00 | 1M | Pricing details → |
| Gemini 3.7 Flash | $0.75 | $0.075 | $3.75 | 1.05M | Pricing details → | |
| Gemini 3.6 FlashScheduled rate change | $0.75 | $0.075 | $3.75 | — | Pricing details → | |
| Gemini 3.5 Flash | $1.50 | $0.15 | $9.00 | 1.05M | Pricing details → | |
| Gemini 3.5 Flash-Lite | $0.30 | $0.030 | $2.50 | — | Pricing details → | |
| Gemini 3.1 Pro PreviewTiered | $2.00–$4.00 | — | $12.00–$18.00 | 1.05M | Pricing details → |
20 models shown
Estimate your LLM API cost
Enter your expected token usage and request volume to estimate monthly TokenGate API spend.
Workload cost calculator
10K requests · 10M input tokens · 5M output tokens
Estimate based on current TokenGate prices. Actual billing follows real token usage.
How AI API pricing works
Input tokens
Tokens sent to the model, including your prompt, instructions and conversation context.
Output tokens
Tokens generated by the model in its response. Output tokens are often priced differently from input tokens.
Cached input
Some models offer a lower rate when previously processed prompt content can be reused from cache.
Tiered pricing
Some models change price based on context size, caching or other request conditions. TokenGate shows those tiers separately instead of reducing them to a misleading single rate.
How to compare LLM API costs
Token price is only one part of model cost. A model with a lower input rate may produce more output tokens, require a larger context or use different pricing for long requests.
When comparing models, use the same expected workload: average input tokens, average output tokens and monthly request volume.
For production selection, compare cost together with model capabilities and context requirements. Use the Models directory for capability-level comparison.
Looking for older model pricing?
TokenGate preserves historical model pages for pricing reference, migration research and replacement guidance.
Pricing FAQ
Choose a model. Estimate the cost. Start building.
Access current AI models through one OpenAI-compatible TokenGate integration and pay in USDT.
