Google

Gemini 3 Flash Preview API

Google's high-speed thinking model designed for agentic workflows, multi-turn chat, and coding assistance. Supports text, image, audio, video, and PDF input with a 1M token context window. Features configurable thinking levels, tool use, and structured output. Broad quality improvements over Gemini 2.5 Flash across reasoning, multimodal understanding, and reliability.

  • Reasoning
  • Tool use
  • Function calling
  • Structured output
  • Long context
Use in Consolegemini-3-flash-preview

USD

Pricing

default
Standard Rate
Input$0.5/1M tokensCompletion Price$3/1M tokensCache Read$0.05/1M tokensCache Write$1/1M tokens

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

API

Code examples

curl --location --request POST 'https://api.tokenhot.ai/v1beta/models/gemini-3-flash-preview:generateContent' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "contents": [
        {
            "role": "user",
            "parts": [
                {
                    "text": "Hello"
                }
        }
    ]
}'

Google

Related models

FAQ

Frequently asked questions

What is Gemini 3 Flash Preview best suited for?

Gemini 3 Flash Preview is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is Gemini 3 Flash Preview priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call Gemini 3 Flash Preview?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with Gemini 3 Flash Preview

Use in Console
TokenHot

The frontier intelligence gateway. One API. 127 models. 0.2s latency. Pay only for what you use.

All systems normal · 99.997% uptime

Company

© 2026 TokenHot Inc. — Built for builders.