Kling O1(Omini One) AI combines images, short videos, and text prompts to provide an integrated video processing solution—from single-frame extension to shot sequences and local reconstruction. It emphasizes consistency and controllability, fitting every creative and production stage.
Compare Kling's supported video models and choose the right generation workflow for your project.
Generate high-quality videos with improved action, audio, and character consistency.
Create cinematic videos with optional sound support, consistent action, and flexible framing.
Turn prompts or images into polished video clips quickly with Kling 2.5 Turbo Pro.
Use a unified multimodal model for text-to-video, image references, and intelligent video editing.
Generate detailed text-to-video scenes with Kling 2.1 Master and stronger control over composition.
Create image-to-video clips with Kling 2.1 Standard for smooth motion and reliable visual consistency.
Kling O1 AI turns complex video manipulation into a process driven by natural language and reference materials. After uploading images or short videos, describe the modifications or generations you want—Kling O1 understands your intent and implements pixel-level changes, enabling creators to focus on concept and expression instead of tedious post-production.
These examples include: generating dynamic shots from still photos, extending footage before or after a reference video, adding interactive props (such as weapons or gifts) to visuals, and overall recoloring or stylization—demonstrating Kling O1's ability to preserve subject integrity and visual consistency.
Focuses on reference-driven subject consistency, semantic-level video editing, shot extension, and style transfer.
Upload multi-angle reference or subject images, and Kling O1 will treat them as one character source, ensuring that appearance, attire, and props remain stable throughout the generation process.
Edit your video using simple text commands—add specific elements, remove unwanted targets, change backgrounds, or shift time of day (like day to dusk).
Generate preceding or following shots, or alternative angles, based on existing footage. Supports description of camera movements (mid-shot, tracking shot, close-up) and pacing to make it easy to build narrative shot sequences.
Combine style transfer and material replacement in one request—apply a specific art style and swap props or costumes as needed for complex creative needs.
Merges high consistency with low-barrier editing, ideal for teams needing repeated content generation and rapid iteration.
Enhances subject feature modeling so that characters maintain a unified appearance across shots, reducing the need for manual fixes.
Natural language serves as the main control panel, reducing toolchain complexity and manual steps, enabling fast multi-round iteration.
A single model can both generate new shots and semantically edit existing videos, reducing cross-tool coordination costs.
Supports referencing both images and videos together to create more creative variations—perfect for advertising and experimental content.
Supports multiple languages for text input and editing, making it easier to work with international teams and content.
Adapted for creative, commercial, and post-production needs.
Quickly generate coherent short shots or storyboard extensions to assist directors and producers with visual validation.
Use product and model images to produce a variety of shooting angles and scene styles, speeding up content creation.
Replace clothing or props in digital samples to produce multiple looks, reducing the cost of physical photography.
Semantically repair errors or unsatisfactory elements in shots, such as replacing backgrounds, removing distractions, or adjusting lighting and color.
Kling O1 is a unified multi-modal video generation and editing engine, supporting images, videos, and text input for an all-in-one creative and editing workflow.