Google

gemini-3.5-flash

Built for multimodal understanding, coding, and long-context tasks, improving over Gemini 3.1 in throughput, context, and tool capabilities. Compared with GPT, Claude, and Qwen models, it is strong for unified analysis workflows across text, images, audio, video, and documents.

  • Reasoning
  • Web search
  • Tool use
  • Function calling
  • Structured output
  • Long context
  • Code execution

AI Chat

Send a prompt and see the response here.

gemini-3.5-flash
Current conversation
How can I help you?

Start a conversation with this model and ask follow-up questions.

Uses your account balance at the model’s current rates.

Insufficient balance

Your account balance is insufficient for this generation. Top up, then return here to try again.

API

Code examples

curl --location --request POST 'https://api.tokenhot.ai/v1beta/models/gemini-3.5-flash:generateContent' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "contents": [
        {
            "role": "user",
            "parts": [
                {
                    "text": "Hello"
                }
        }
    ]
}'

FAQ

Frequently asked questions

What is gemini-3.5-flash best suited for?

gemini-3.5-flash is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is gemini-3.5-flash priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call gemini-3.5-flash?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with gemini-3.5-flash

Use in Console