Google

Gemini 3.1 PRO Preview API

Built for multimodal understanding, coding, and long-context tasks, improving over Gemini 3.0 / 2.5 in throughput, context, and tool capabilities. Compared with GPT, Claude, and Qwen models, it is strong for unified analysis workflows across text, images, audio, video, and documents.

  • Long context
Use in Consolegemini-3.1-pro-preview

USD

Pricing

default
input<20K
Input$2/1M tokensCompletion Price$12/1M tokensCache Read$0.014/1M tokensCache Write$4.5/1M tokens
input>=27K
Input$4/1M tokensCompletion Price$18/1M tokensCache Read$0.014/1M tokensCache Write$4.5/1M tokens

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

API

Code examples

curl --location --request POST 'https://api.tokenhot.ai/v1beta/models/gemini-3.1-pro-preview:generateContent' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "contents": [
        {
            "role": "user",
            "parts": [
                {
                    "text": "Hello"
                }
        }
    ]
}'

Google

Related models

FAQ

Frequently asked questions

What is Gemini 3.1 PRO Preview best suited for?

Gemini 3.1 PRO Preview is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is Gemini 3.1 PRO Preview priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call Gemini 3.1 PRO Preview?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with Gemini 3.1 PRO Preview

Use in Console
TokenHot

The frontier intelligence gateway. One API. 127 models. 0.2s latency. Pay only for what you use.

All systems normal · 99.997% uptime

Company

© 2026 TokenHot Inc. — Built for builders.