Gemini 3 Flash Preview
Google's high-speed thinking model designed for agentic workflows, multi-turn chat, and coding assistance. Supports text, image, audio, video, and PDF input with a 1M token context window. Features configurable thinking levels, tool use, and structured output. Broad quality improvements over Gemini 2.5 Flash across reasoning, multimodal understanding, and reliability.
- Reasoning
- Tool use
- Function calling
- Structured output
- Long context
AI Chat
Send a prompt and see the response here.
gemini-3-flash-previewStart a conversation with this model and ask follow-up questions.
USD /1M tokens
Pricing
| Group | Input | Output | Cache Read | Cache Write |
|---|---|---|---|---|
| default | $0.5 | $3 | $0.05 | $1 |
Public list prices are shown in USD. Final charges may vary by account group and usage tier.
API
Code examples
curl --location --request POST 'https://api.tokenhot.ai/v1beta/models/gemini-3-flash-preview:generateContent' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
"contents": [
{
"role": "user",
"parts": [
{
"text": "Hello"
}
}
]
}'FAQ
Frequently asked questions
What is Gemini 3 Flash Preview best suited for?
Gemini 3 Flash Preview is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.
How is Gemini 3 Flash Preview priced?
Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.
How can I call Gemini 3 Flash Preview?
Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.
Get started
Build with Gemini 3 Flash Preview
API