Gemini 3.1 Pro Preview API pricing
Google's Pro preview model for complex reasoning, coding and multimodal agentic work.
How much does Gemini 3.1 Pro Preview API cost?
Gemini 3.1 Pro Preview uses variable TokenGate pricing. The rates below depend on the request conditions listed in each row.
| Condition | Input tokens / 1M | Cached input / 1M | Output tokens / 1M |
|---|---|---|---|
| Prompts ≤ 200K tokens | $2.00 | — | $12.00 |
| Prompts > 200K tokens | $4.00 | — | $18.00 |
- Pricing is tiered by prompt size.
Use Gemini 3.1 Pro Preview with the OpenAI SDK
curl https://api.tokengate.com/v1/chat/completions \
-H "Authorization: Bearer $TOKENGATE_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"gemini-3.1-pro-preview","messages":[{"role":"user","content":"Hello"}]}'Change the model identifier to use another supported model through the same TokenGate integration.
Model specifications
- Provider
- API model identifier
- gemini-3.1-pro-preview
- Generation
- 3.1
- Provider lifecycle
- Active
- Context window
- 1.05M
- Maximum output tokens
- 66K
- Input modalities
- Text, Image
- Output modalities
- Text
- Capabilities
- Streaming, Function calling, Structured outputs, Vision, Reasoning, Tool use
Best for
- Complex reasoning
- Coding
- Agents
- Multimodal
Access Gemini 3.1 Pro Preview through TokenGate
Use supported models through a single TokenGate integration.
- · One API integration
- · OpenAI-compatible interface
- · Payment in USDT
Frequently asked questions
Related models
DeepSeek V4.1 Flash
DeepSeek's current Flash model with native multimodal support and strong coding, reasoning and agentic capabilities.
- Input / 1M
- Variable pricing
- Output / 1M
- —
- Context
- —
- Generation
- 4.1
Streaming · Function calling · Structured outputs
DeepSeek V4.1 Flash API pricingGemini 3.7 Flash
Previous-generation Google Flash model with multimodal input and a large context window.
- Input / 1M
- $0.75
- Output / 1M
- $3.75
- Context
- 1.05M
- Generation
- 3.7
Streaming · Function calling · Structured outputs
Gemini 3.7 Flash API pricingGemini 3.5 Flash
Google Flash model with a large context window for high-volume multimodal workloads.
- Input / 1M
- $1.50
- Output / 1M
- $9.00
- Context
- 1.05M
- Generation
- 3.5
Streaming · Function calling · Structured outputs
Gemini 3.5 Flash API pricingTokenGate
