LLM API Pricing

DeepSeek Models

Browse all 4 DeepSeek models available on TokenHot.

One API key gives you access to 4 DeepSeek models with live pricing and no separate provider account.

DeepSeek
CONTEXT1M
Input$0.3000/M
Output$1.2000/M
A strong high-throughput, lower-cost default for general and agentic work: it natively understands images and supports a 1M-token context, optional thinking and non-thinking modes, tool calls, and JSON output. Compared with the higher-tier V4 Pro, it is positioned around speed, throughput, and cost efficiency, making it practical for long documents, coding/agents, and image-text analysis; enable thinking when deeper reasoning is needed.
Input Type:
Output Type:
ReasoningTool UseStructured OutputLong Context
Try model
CONTEXT1.05M
Input$0.1490/M
Output$0.2980/M
An experimental multimodal variant of DeepSeek-V4-Flash for image understanding alongside text. A practical choice for visual Q&A, screenshot and document analysis, while its experimental status makes it better suited to evaluation and flexible workflows than strict production-critical paths.
Input Type:
Output Type:
ReasoningTool UseLong Context
Try model
DeepSeek
CONTEXT1M
Input
$1.5390/M$1.7100/M
Output
$3.0869/M$3.4299/M
Built for reasoning, coding, and agentic tool use, with stronger long-context, thinking-mode, and complex-task execution than DeepSeek V3.2. Compared with Chinese peers such as Qwen, GLM, and Kimi, it is especially competitive for math, programming, and logical reasoning tasks.
Input Type:
Output Type:
Long ContextTool UseReasoningStructured Output
Try model
DeepSeek
CONTEXT1M
Input
$0.1278/M$0.1420/M
Output
$0.2574/M$0.2860/M
Built for reasoning, coding, and agentic tool use, with stronger long-context, thinking-mode, and complex-task execution than DeepSeek V3.2. Compared with Chinese peers such as Qwen, GLM, and Kimi, it is especially competitive for math, programming, and logical reasoning tasks.
Input Type:
Output Type:
Long ContextTool UseReasoningStructured Output
Try model