Kling 3.0 is Kuaishou's video generation model, released February 6, 2026. It merged two earlier lines onto a unified multimodal framework: Kling VIDEO 2.6 became Kling 3.0, and Kling VIDEO O1 became Kling 3.0 Omni. It generates 3 to 15 seconds of video from text, images, video or audio input, with sound rendered alongside the picture rather than added afterward. The capabilities that arrived with it and were absent from Kling 2.6 are multi-shot narratives, element reference, multi-character coreference across three or more speakers, five languages with dialects and accents, and flexible duration up to 15 seconds.