On September 28, Kling AI officially announced that the newly upgraded

Comprehensive audio-visual performance breakthroughs, approaching cinematic quality
Kling4.0 significantly improves stability and continuity when dealing with large-scale camera movements such as high-speed motion, continuous action, tracking shots, and surrounding shots. In terms of audio, it now supports high-quality audio and dual-channel stereo sound, greatly enhancing the accuracy of character lip-syncing. In terms of video output, it not only supports ultra-wide 21:9 aspect ratio and 4K resolution, but also the 10-bit HDR format at 1080p will be released soon.
Extreme multi-modal control, reshaping commercial video creation
In terms of creative control, the new version brings unprecedented freedom: it supports up to 15 multi-modal reference inputs per generation (including 10 images, 5 videos, and 7 subject combinations). It not only allows precise editing of the subject and background of the original video, but can also intelligently analyze the cinematography and narrative structure of the reference video, making it easy to generate new commercial-level videos.
Strong long-take storytelling and multi-language support
Kling4.0 now supports generating native videos up to 30 seconds in a single session, perfectly meeting the needs of long takes and continuous storytelling. With up to 10 keyframe inputs, creators can precisely control the characters' states and plot development. Additionally, the prompt input limit has been significantly expanded to 8000 Tokens, supporting multiple languages including Chinese, English, Japanese, Korean, Spanish, Portuguese, German, French, and Hindi, as well as various accents and dialects. The ability to extend videos up to 2 minutes through multiple continuations will also be released soon.
Completely renewed workflow, delivering an ultimate user experience
Join Now