LLM API Pricing

ByteDance Models

Browse all 8 ByteDance models available on TokenHot.

One API key gives you access to 8 ByteDance models with live pricing and no separate provider account.

Input
$71.2500/M$75.0000/M
Output
$71.2500/M$75.0000/M
A better fit for moving from a single clip toward a complete audiovisual story: it can generate up to 30 seconds in one pass, supports multiple extensions, accepts image, video, and audio references, and adds timestamp-level editing, green-screen, and camera control. It suits ads and short films that need continuity, while complex multi-subject motion still deserves human review.
Input Type:
Output Type:
Multimodal Output
Try model
Input$3.3042/M
Output$3.3042/M
A route variant of Seedance 2.5; evaluate it by the underlying model rather than the route suffix. It is suited to ads, branded content, and narrative shorts, with up to 30-second audio-video generation, multi-round extension, text/image/audio/video references, R2V control, and timestamp-level editing. Complex motion and multi-subject physical consistency still benefit from human review.
Input Type:
Output Type:
Multimodal Output
Try model
Input$0.6300/M
Output$0.6300/M
Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. This filter-off variant keeps the same tier of generation capability as standard Seedance 2.0 while applying looser content filtering. Useful for short drama, ads, e-commerce, and social assets.
Input Type:
Output Type:
Multimodal Output
Try model
CONTEXT13K
Input$2.1606/M
Output$2.1606/M
Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. This filter-off variant keeps the same tier of generation capability as standard Seedance 2.0 while applying looser content filtering. Useful for short drama, ads, e-commerce, and social assets.
Input Type:
Output Type:
Multimodal Output
Try model
CONTEXT13K
Input$1.5465/M
Output$1.5465/M
Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. This filter-off variant keeps the same tier of generation capability as standard Seedance 2.0 while applying looser content filtering. Useful for short drama, ads, e-commerce, and social assets.
Input Type:
Output Type:
Multimodal Output
Try model
CONTEXT8K
Input
$1.3145/M$1.5465/M
Output
$1.3145/M$1.5465/M
Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. The fast variant is better for rapid iteration and batch production than the standard tier. Useful for short drama, ads, e-commerce, and social assets.
Input Type:
Output Type:
Multimodal Output
Try model
CONTEXT8K
Input
$2.0526/M$2.1606/M
Output
$2.0526/M$2.1606/M
Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. Useful for short drama, ads, e-commerce, and social assets.
Input Type:
Output Type:
Multimodal Output
Try model
PER SEC
$0.0125/s$0.0250/s
A lighter, throughput-oriented variant in the Seedance 2.0 family for short-video drafts, batch assets, and cost-sensitive iteration. Its value is the lighter generation tier; choose the standard model when maximum fidelity, complex motion, or higher output specifications matter more.
Input Type:
Output Type:
Multimodal Output
Try model