Kling 3.0 Omni
Locks character look and voice across multi-shot scenes.
Professional video for every use case
Recurring character shorts.
Reference images and bound voices keep the same lead believable across episodes, camera angles, and locations.Talking-head explainers.
Lip-synced speech and ambient audio make short scripted lessons feel finished without separate voiceover assembly.Multilingual social ads.
Localized dialogue and synced mouth movement help you adapt one concept for different audiences faster.Previs and storyboards.
Shot-by-shot prompts turn scripts into paced 15-second scenes you can review before filming.Reference images, elements, and video together
Kling 3.0 Omni is built for reference-driven generation. You can combine text with images, reusable elements, and a short video input so the model stays anchored to a real subject instead of reinventing it each time.
- •Text-to-video, image-to-video, and start/end frame workflows
- •Upload up to 7 images or elements when no video is attached, or use one 3-10 second video plus up to 4 images or elements
- •Mix product, character, and scene references in one prompt
- •Reference-driven input reduces drift
Voice-bound characters with native audio
You can create a reusable character from multi-angle images or a short character video, then bind a voice to that subject. The model generates dialogue, ambient sound, and lip-sync together, which makes recurring characters far easier to reuse.
- •Bind voice from a 5-30 second speech clip, or extract look and voice from a 3-8 second character video
- •Generate dialogue, ambience, and sound effects together
- •Supports English, Chinese, Japanese, Korean, and Spanish
- •Built for recurring characters and talking scenes
Shot-by-shot control up to 15 seconds
Instead of stitching separate clips, you can describe multiple beats inside one generation. Kling 3.0 Omni lets you set shot length, framing, angle, narrative action, and camera movement while keeping continuity across the sequence.
- •Flexible duration from 3 to 15 seconds
- •Plan up to 6 cuts in one generation
- •Set framing, angle, and camera movement per beat
- •Smooth transitions between shots
- •Useful for dialogue, reveals, and micro-stories
Start/end frames and video editing tools
Kling says Omni carries over the frame guidance, editing, and prompt transformation tools from its earlier multimodal workflow. That gives you a practical way to extend, restyle, or selectively change shots instead of starting every revision from scratch.
- •Use start and end frames to guide motion
- •Generate previous or next shots from a reference video
- •Change subjects, backgrounds, or selected details
- •Restyle footage with prompt-based edits
- •Useful for revisions and continuity fixes
How it works
Add a prompt or references
Start with a text prompt, a product image, multi-angle character photos, or a short reference video. If consistency matters, upload the subject you want preserved instead of relying on text alone.
Bind voice and choose settings
For speaking scenes, add a clean voice clip or a short character video so the model can lock both look and voice. Then choose duration, resolution, audio mode, and whether the scene should play as one shot or several beats.
Generate and refine
Review the first pass for identity, lip-sync, and shot timing. If something drifts, tighten the shot directions, strengthen your references, or guide motion with start and end frames.
Pricing for Kling 3.0 Omni
Runs on credits — no per-model surcharges, no surprise billing.
Use Kling V3 Omni via the API
Kling V3 Omni is available through the BudgetPixel developer API — the same model the studio runs, supporting text-to-video, image-to-video, reference-to-video, and video-to-video. Pricing is metered in credits (85 credits/second at 720p, 115 credits/second at 1080p, 450 credits/second at 4k), charged only on success, with an API key available on Premium plans and above.
curl -X POST https://api.budgetpixel.com/v1/videos/kling-v3-omni-video \
-H "Authorization: Bearer $BUDGETPIXEL_API_KEY" \
-H "Content-Type: application/json" \
-d '{"prompt": "your prompt here"}'