Gemini 3 Flash Preview API
Google's high-speed thinking model designed for agentic workflows, multi-turn chat, and coding assistance. Supports text, image, audio, video, and PDF input with a 1M token context window. Features configurable thinking levels, tool use, and structured output. Broad quality improvements over Gemini 2.5 Flash across reasoning, multimodal understanding, and reliability.
- Reasoning
- Tool use
- Function calling
- Structured output
- Long context
gemini-3-flash-previewUSD
Pricing
Public list prices are shown in USD. Final charges may vary by account group and usage tier.
API
Code examples
curl --location --request POST 'https://api.tokenhot.ai/v1beta/models/gemini-3-flash-preview:generateContent' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
"contents": [
{
"role": "user",
"parts": [
{
"text": "Hello"
}
}
]
}'Related models
FAQ
Frequently asked questions
What is Gemini 3 Flash Preview best suited for?
Gemini 3 Flash Preview is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.
How is Gemini 3 Flash Preview priced?
Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.
How can I call Gemini 3 Flash Preview?
Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.
Get started