LLM API Pricing

Qwen Models

Browse all 15 Qwen models available on TokenHot.

One API key gives you access to 15 Qwen models with live pricing and no separate provider account.

Qwen
qwen3.8-max
Qwen
CONTEXT1M
Input
$1.4200/M$1.7750/M
Output
$4.2640/M$5.3300/M
Qwen's most capable flagship multimodal reasoning model, built for long-horizon coding, professional work, and multi-stage agent tasks. It accepts text, images, and video, offers a 1M-token context window with up to 128K output, and supports function calling, structured outputs, web search, and a code interpreter. It is best used as a high-quality workhorse for complex tasks rather than for the lowest-cost, lowest-latency batch workloads.
Input Type:
Output Type:
ReasoningWeb SearchTool UseFunction CallingStructured OutputLong ContextCode Execution
Qwen
happyhorse-1.1-t2v
阿里巴巴
PER SEC$0.1350/s
Built for text-to-video, with a stronger audio-native workflow than HappyHorse 1.0, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Multimodal Output
Qwen
happyhorse-1.1-r2v
阿里巴巴
PER SEC$0.1350/s
Built for reference-to-video, with a stronger audio-native workflow than HappyHorse 1.0, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Multimodal Output
Qwen
happyhorse-1.1-i2v
阿里巴巴
PER SEC$0.1350/s
Built for image-to-video, with a stronger audio-native workflow than HappyHorse 1.0, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Multimodal Output
Qwen
qwen3.7-max
Qwen
CONTEXT1M
Input
$1.4400/M$1.8000/M
Output
$4.3200/M$5.4000/M
Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.6. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Qwen
qwen3.6-max-preview
Qwen
CONTEXT262K
Input$1.3300/M
Output$8.0000/M
Alibaba Qwen 3.6 Max Preview, a sparse mixture-of-experts model with ~1 trillion parameters. Optimized for agentic coding, tool use, and long-context reasoning with an integrated thinking mode that preserves reasoning traces across multi-turn conversations. 262K context window, available via Alibaba Cloud Model Studio API.
Input Type:
Output Type:
ReasoningTool UseStructured OutputLong Context
Qwen
happyhorse-1.0-video-edit
Qwen
CONTEXT8K
PER SEC$0.1350/s
Built for video editing, with a stronger audio-native workflow than 早期视频模型, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Qwen
happyhorse-1.0-t2v
Qwen
CONTEXT8K
PER SEC$0.1350/s
Built for text-to-video, with a stronger audio-native workflow than 早期视频模型, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Qwen
happyhorse-1.0-r2v
Qwen
CONTEXT8K
PER SEC$0.1350/s
Built for reference-to-video, with a stronger audio-native workflow than 早期视频模型, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Qwen
happyhorse-1.0-i2v
Qwen
CONTEXT8K
PER SEC$0.1350/s
Built for image-to-video, with a stronger audio-native workflow than 早期视频模型, generating dialogue, sound effects, and background music in one pass. Compared with Veo, Kling, and Runway, its differentiator is synchronized audiovisual generation and character-consistent workflows for short drama, ads, e-commerce, and brand marketing.
Input Type:
Output Type:
Qwen
qwen3.7-plus
Qwen
CONTEXT1M
Input
$0.2400/M$0.3000/M
Output
$0.9600/M$1.2000/M
Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.6. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Qwen
qwen3.6-plus
Qwen
CONTEXT1M
Input
$0.9120/M$1.1400/M
Output
$5.4880/M$6.8599/M
Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.5. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.
Input Type:
Output Type:
ReasoningTool UseStructured OutputLong Context
Qwen
qwen3.6-flash
Qwen
CONTEXT1M
Input$0.1785/M
Output$1.0605/M
Designed for agentic workloads, coding, and office productivity, with a stronger combined focus on long context, tool calling, and structured output than Qwen3.5. Compared with flagship models such as GPT, Claude, and Gemini, it is a cost-effective alternative for Chinese, coding, and multi-step tool workflows.
Input Type:
Output Type:
ReasoningTool UseStructured OutputLong Context
Qwen
qwen-image-2.0-pro
Qwen
Input$2.2000/M
Output$2.2000/M
Built for image generation and editing, improving over Qwen Image 1.x in Chinese text rendering, instruction following, and complex layouts. Compared with GPT Image, Gemini image models, and Seedream, it is well suited to Chinese posters, e-commerce images, infographics, and iterative visual creation.
Input Type:
Output Type:
Multimodal Output
Qwen
qwen-image-2.0
Qwen
Input$2.2000/M
Output$2.2000/M
Built for image generation and editing, improving over Qwen Image 1.x in Chinese text rendering, instruction following, and complex layouts. Compared with GPT Image, Gemini image models, and Seedream, it is well suited to Chinese posters, e-commerce images, infographics, and iterative visual creation.
Input Type:
Output Type:
Multimodal Output