Explore Kuaishou's Kling AI family—from text and image-led video to native audio, motion control and unified multimodal creation in the 3.0 series.
Kling overview
Kuaishou's generative media platform for video, image, sound and effects.
Kling video models support prompt-directed motion, image animation and increasingly complete production workflows that combine understanding, generation and editing.
The 3.0 series moves toward a native multimodal architecture. Available features differ across VIDEO 3.0, Omni and Motion Control products and must not be combined into one generic endpoint claim.
Capabilities
Specific availability varies by generation and product.
Evolution guide
Each generation expands the production workflow.
| Generation | Family role | Page-safe summary |
|---|---|---|
| Kling 1.x | Foundation | Established text-to-video, image-to-video and controllable motion workflows. |
| Kling 2.x | Production expansion | Improved quality and control; 2.6 added simultaneous audio-video and motion control. |
| Kling 3.0 | Native multimodal series | Combines understanding, generation and editing across text, image, audio and video. |
Workflows
Use cases
Kling FAQ
Understand the standard, Omni and Motion Control boundaries before access opens.