Kling Image O1
[Core Function] Kling Image O1 is a reasoning-enhanced multimodal image model. [Strengths] It performs deep reasoning over prompts and references to handle complex logic, spatial relationships, and intricate multi-image combinations. [Best For] Highly recommended for: complex scenes requiring strict logical or spatial accuracy. [Limitations] Element library IDs and series generation are not exposed. Default aspect_ratio is 1:1 when omitted. Do NOT use for simple artistic generation where V3 is faster and more stylistic. [Routing] Route to this model when the prompt involves complex physical logic or strict spatial reasoning.
Authorizations
API Key authentication. Format: Bearer YOUR_API_KEY.
Body
Kling Image O1 (POST /images/omni-image, model_name=kling-image-o1). No result_type/series_amount. Element is not exposed. Default aspect_ratio is 1:1 when omitted.
Text prompt (positive/negative in one field). May reference images as <<<image_1>>> etc. Max 2500 characters.
1 - 2500"Merge all the people into the <<<image_1>>> scene"
Optional reference images (flattened image_list). Omit for pure text-to-image. If provided, must contain 1-10 non-empty URLs/Base64 strings.
1 - 10 elementsImage URL or Base64 (JPG/JPEG/PNG, ≤10MB, ≥300px, aspect 1:2.5~2.5:1)
1Output resolution. Capability Map for O1: 1k | 2k only (no 4k).
1k, 2k "2k"
Aspect ratio. Default 1:1 when omitted (filled by platform).
16:9, 9:16, 1:1, 4:3, 3:4, 3:2, 2:3, 21:9 "1:1"
Number of images to generate [1, 9].
1 <= x <= 91