kling

V3 API

Built for text/image-to-video and multi-shot creation, improving over Kling 2.x in motion stability, camera language, and reference consistency. Compared with Veo, Runway, and Seedance, it is well suited to Chinese short drama, ad assets, and high-quality social video production.

  • Multimodal output

USD

Pricing

default
720P$0.09/s
有声,4K$0.09/s
无声,1080P$0.1197/s
有声,720P$0.135/s
有声,1080P$0.135/s
无声,4K$0.45/s

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

API

Code examples

curl --location 'https://api.tokenhot.ai/v1/video/generations' \
--header 'Content-Type: application/json' \
--data '{
    "model": "kling-v3",
    "prompt": "A white long-haired cat lounges on a wooden windowsill in the sun, blooming roses in the background; the cat blinks slowly and gently flicks its tail",
    "duration": 5,
    "size": "1080P",
    "aspect_ratio": "16:9",
    "audio_generation": true
}'

kling

Related models

KlingklingV3 Omni

Built for text/image-to-video and multi-shot creation, improving over Kling 2.x in motion stability, camera language, and reference consistency. Compared with Veo, Runway, and Seedance, it is well suited to Chinese short drama, ad assets, and high-quality social video production.

$0.09/s
DoubaoDoubaoSeedance 2 0 Fast Filter OFF

Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. This filter-off variant keeps the same tier of generation capability as standard Seedance 2.0 while applying looser content filtering. Useful for short drama, ads, e-commerce, and social assets.

$4.2/1M tokens
DoubaoDoubaoSeedance 2 0

Built for video generation and multi-asset video creation, improving on Seedance 1.x with more complete motion stability, audiovisual generation, and reference inputs. Compared with Veo, Kling, and Runway, it is better suited to Chinese creative workflows driven by mixed text, image, audio, and video inputs. Useful for short drama, ads, e-commerce, and social assets.

$4.788/1M tokens
GeminiGoogleGemini Omni Video

Built for any-input-to-video multimodal creation, putting stronger emphasis than earlier Gemini video workflows on unified orchestration of text, image, audio, and video materials. Compared with Veo, Kling, and Runway, it is better for integrating complex multimodal assets into one creative workflow.

$0.495/s

FAQ

Frequently asked questions

What is V3 best suited for?

V3 is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is V3 priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call V3?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with V3

Use in Console