LLM API Pricing

Google Models

Browse all 15 Google models available on TokenHot.

One API key gives you access to 15 Google models with live pricing and no separate provider account.

Gemini
gemini-omni-video
Google
PER SEC$0.2450/s
Built for any-input-to-video multimodal creation, putting stronger emphasis than earlier Gemini video workflows on unified orchestration of text, image, audio, and video materials. Compared with Veo, Kling, and Runway, it is better for integrating complex multimodal assets into one creative workflow.
Input Type:
Output Type:
Multimodal Output
Gemini
gemini-3.5-flash
Google
CONTEXT1M
Input$1.5000/M
Output$9.0000/M
Built for multimodal understanding, coding, and long-context tasks, improving over Gemini 3.1 in throughput, context, and tool capabilities. Compared with GPT, Claude, and Qwen models, it is strong for unified analysis workflows across text, images, audio, video, and documents.
Input Type:
Output Type:
ReasoningWeb SearchTool UseFunction CallingStructured OutputLong ContextCode Execution
Gemini
gemini-3.1-flash-image-preview
Google
CONTEXT131K
PER IMG$0.0672/image
Built for image generation and editing, with stronger multimodal understanding, conversational edits, and rapid visual prototyping than earlier Gemini image capabilities. Compared with GPT Image, Qwen Image, and Seedream, it is useful for combining documents, images, and prompts in one visual workflow.
Input Type:
Output Type:
ReasoningMultimodal Output
Gemini
gemini-3.1-pro-preview
Google
CONTEXT1.05M
Input$2.0000/M
Output$12.0000/M
Built for multimodal understanding, coding, and long-context tasks, improving over Gemini 3.0 / 2.5 in throughput, context, and tool capabilities. Compared with GPT, Claude, and Qwen models, it is strong for unified analysis workflows across text, images, audio, video, and documents.
Input Type:
Output Type:
Long Context
Gemini
gemini-3-flash-preview
Google
CONTEXT1M
Input$0.5000/M
Output$3.0000/M
Google's high-speed thinking model designed for agentic workflows, multi-turn chat, and coding assistance. Supports text, image, audio, video, and PDF input with a 1M token context window. Features configurable thinking levels, tool use, and structured output. Broad quality improvements over Gemini 2.5 Flash across reasoning, multimodal understanding, and reliability.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Gemini
gemini-3-pro-image-preview
Google
CONTEXT66K
PER IMG$0.1340/image
Built for image generation and editing, with stronger multimodal understanding, conversational edits, and rapid visual prototyping than earlier Gemini image capabilities. Compared with GPT Image, Qwen Image, and Seedream, it is useful for combining documents, images, and prompts in one visual workflow.
Input Type:
Output Type:
ReasoningMultimodal Output
Gemini
veo3.1-lite
Google
PER CALL$0.1100/call
Built for high-fidelity video generation, with a more mature native-audio, camera-control, and reference-input workflow than Veo 2/3. Compared with Kling, Runway, and Seedance, it is better suited to cinematic shorts, ad storyboards, and videos requiring stable physical motion. The Lite tier favors high-throughput, lower-cost drafts over the standard tier.
Input Type:
Output Type:
Multimodal Output
Gemini
veo3.1-fast
Google
PER CALL$0.2211/call
Built for high-fidelity video generation, with a more mature native-audio, camera-control, and reference-input workflow than Veo 2/3. Compared with Kling, Runway, and Seedance, it is better suited to cinematic shorts, ad storyboards, and videos requiring stable physical motion. The Fast tier is better for rapid concept validation than the standard tier.
Input Type:
Output Type:
Multimodal Output
Gemini
veo3.1
Google
PER CALL$1.6544/call
Built for high-fidelity video generation, with a more mature native-audio, camera-control, and reference-input workflow than Veo 2/3. Compared with Kling, Runway, and Seedance, it is better suited to cinematic shorts, ad storyboards, and videos requiring stable physical motion.
Input Type:
Output Type:
Multimodal Output
Gemini
gemini-2.5-flash-image
Google
CONTEXT66K
PER IMG$0.0585/image
Built for image generation and editing, with stronger multimodal understanding, conversational edits, and rapid visual prototyping than earlier Gemini image capabilities. Compared with GPT Image, Qwen Image, and Seedream, it is useful for combining documents, images, and prompts in one visual workflow.
Input Type:
Output Type:
Structured OutputMultimodal Output
Gemini
gemini-3.6-flash
Google
Input$1.5000/M
Output$7.5000/M
Gemini
gemini-3.5-flash-lite
Google
Input$0.3000/M
Output$2.5000/M
Gemini
gemini-3.1-flash-lite-image
Google
PER CALL$0.0200/call
Gemini
gemini-3.1-flash-image
Google
PER IMG$0.0330/image
Gemini
gemini-3-pro-image
Google
PER IMG$0.0530/image