The Kling 3.0 series models API is now fully available
Learn More
Get Started
Overview
Quick Start
Changelog
API Reference
General Info
Rate Limits
Callback Schema
Video Generation
Models
Video Omni
Text to Video
Image to Video
Reference to Video
Motion Control
Multi-elements to video
Extend Video
Lip Sync
Avatar
Text to Audio
Video to Audio
Text to Speech
Voice Clone
Image Recognize
Element
Effects
Effect Templates
NEW
Video Effects
Image Generation
Models
Image Omni
Image Generation
Reference to Image
Extend Image
AI Multi-Shot
Virtual Try-On
Others
Query user info
Pricing
Billing Info
Prepaid Resource Packs
Protocols
Privacy Policy of API Service
Terms of API Service
API Service Level Agreement
Image Models

kling-image-o1

	
custom aspect ratio（1K/2K）

	
intelligent aspect ratio


text to image    

	
single-image generation

	
✅

	
✅


others

	
-

	
-


image to image

	
single-image generation

	
✅

	
✅


element control

（only multi-image elements）

	
✅

	
✅


others

	
-

	
-

kling-v3-omni

	
custom aspect ratio（1K/2K/4K）

	
intelligent aspect ratio


text to image

	
single-image generation

	
✅

	
✅


others

	
-

	
-


image to image

	
single-image generation

	
✅

	
✅


series-image generation

	
✅

	
✅


element control

（only multi-image elements）

	
✅

	
✅


others

	
-

	
-

kling-v1

	
1:1

	
16:9

	
4:3

	
3:2

	
2:3

	
3:4

	
9:16

	
21:9


text to image

	
-

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
-


image to image

	
entire image

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
-


others

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-

kling-v1-5

	
1:1

	
16:9

	
4:3

	
3:2

	
2:3

	
3:4

	
9:16

	
21:9


text to image

	
-

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


image to image

	
subject

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


face

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


others

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-

kling-v2

	
1:1

	
16:9

	
4:3

	
3:2

	
2:3

	
3:4

	
9:16

	
21:9


text to image

	
-

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


image to image

	
multi-image to image

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


restyle

	
✅  (The resolution of the generated image is the same as that of the input image, and it does not support setting the resolution separately)


others

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-

kling-v2-new

	
1:1

	
16:9

	
4:3

	
3:2

	
2:3

	
3:4

	
9:16

	
21:9


text to image

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-


image to image

	
restyle

	
✅  (The resolution of the generated image is the same as that of the input image, and it does not support setting the resolution separately)


others

	
-

	
-

	
-

	
-

	
-

	
-

	
-

	
-

kling-v2-1

	
1:1

	
16:9

	
4:3

	
3:2

	
2:3

	
3:4

	
9:16

	
21:9


text to image

	
-

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


image to image

	
entire image

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
-


subject

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


face

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


multi-image to image

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅

	
✅


restyle

	
✅  (The resolution of the generated image is the same as that of the input image, and it does not support setting the resolution separately)

kling-v3

	
custom aspect ratio（1K/2K）

	
intelligent aspect ratio


text to image    

	
single-image generation

	
✅

	
-


others

	
-

	
-


image to image

	
single-image generation

	
✅

	
-


element control

（only multi-image elements）

	
✅

	
-


others

	
-

	
-

no related of model

	
support or not

	
description


image expansion

	
✅

	
Supports expand content based on existing images


others

	
-

	
Model

	
kling-v1

	
kling-v1-5

	
kling-2


Feature

	
Text to Image

	
Image to Image

	
Text to Image

	
Image to Image

	
Text to Image

	
Image to Image


Resolution

	
1K

	
1K

	
1K

	
1K

	
1K/2K

	
1K

Previous chapter：Video Effects
Next chapter：Image Omni
The Kling 3.0 Series Models API is Now Fully Available
– All in One, One for All！

Models Available in This Release

Kling 3.0 Motion Control, Kling Video 3.0, Kling Video 3.0 Omni, Kling Image 3.0, Kling Image 3.0 Omni

Refer to <Kling AI Series 3.0 Model API Specification>

Key Highlights of the Models

3.0 All-in-One: A unified model for multi-modal input and output.

Most powerful consistency across the universe: Subject consistency (supports cameo, subject with voice control, i2v + subject) and text consistency.
Narrative control at your fingertips: More freedom, precision, and control—up to 15 seconds long, video scene cuts, ultra-high-definition storyboards/images, custom seconds.
Upgraded native audio-visual output: Supports multiple speakers and languages (with accents).

Kling 3.0 Motion Control

Consistent Facial Identity from any angle
Complex Emotions faithfully reproduced
High fidelity Restoration, Even with Face Occlusions
Consistent Facial Clarity Across Dynamic Framing

User Guide ->

Kling Video 3.0

Compared to 2.6, expected improvements:

Supports subject upload in I2V scenarios for enhanced consistency
Significant improvement in multi-character referencing, especially for three-person scenarios
Supports Japanese, Korean, and Spanish in addition to Chinese and English
Capable of generating certain dialects and accents
Better distinction and control over different types of audio (speech, sound effects, BGM)
Improved text retention in I2V scenarios
Supports scene transitions, with up to 6 shots and customizable storyboarding

User Guide ->

Kling Video 3.0 Omni

Compared to O1, expected improvements:

Native audio-visual synchronization
Supports video subject creation
Further improved consistency in reference-based tasks, especially for characters and products
Combined capabilities of reference + storyboarding + audio-visual sync significantly enhance usability
Supports scene transitions, with up to 6 shots
Extended generation duration up to 15 seconds

User Guide ->

Kling Image 3.0

Highly consistent feature retention
Precise response to detail modifications
Accurate control over style and tone
Rich imaginative capabilities

User Guide ->

Kling Image 3.0 Omni

Enhanced narrative sense
New storyboard image set generation, retaining reference image features with scene relevance
Direct output of 2K/4K ultra-high-definition images
Further improved detail consistency

User Guide ->

Thank you for your support and understanding!

I Got It