# Kling O1 Image 2 Video > Transforms static images into dynamic, physics-driven animations for creative storytelling. ## Overview - **Endpoint**: `https://api.segmind.com/v1/kling-o1-image-to-video` - **Model ID**: `kling-o1-image-to-video` - **Category**: Image-to-Video Generation - **Type**: Synchronous (Direct response) - **Average Latency**: ~75.8s (30-day average) - **Average Cost**: $0.7056 per run (observed across past runs, not a price — see Pricing) - **Provider**: EACHLABS ## Pricing Usage-based pricing: | Metric | Rate | | --- | --- | | Output Duration | $0.140 per second | ## API Information This model uses a **synchronous response pattern**: 1. Make a POST request with your parameters 2. Receive the output directly in the response (binary for images/videos/audio, JSON for text) 3. No polling required - response is immediate ### Input Schema The API accepts the following input parameters: - **`prompt`** (`string`, _required_): Prompt Describe the video transformation or animation between start and end images - **`start_image_url`** (`File (URL)`, _required_): Start image uRL URL of the starting frame image - **`end_image_url`** (`File (URL)`, _optional_): End image uRL URL of the ending frame image (optional) - **`duration`** (`integer | null`, _optional_): Duration Video duration in seconds - Default: `5` - Options: "5" (5), "10" (10) **Required Parameters Example**: ```json { "prompt": "A seamless transformation of a man driving a vehicle and crashing", "start_image_url": "https://segmind-resources.s3.amazonaws.com/output/d472b3e1-2b68-406d-9d12-a690c3da4045-seedance_1.5_input.webp" } ``` **Full Example**: ```json { "prompt": "A seamless transformation of a man driving a vehicle and crashing", "start_image_url": "https://segmind-resources.s3.amazonaws.com/output/d472b3e1-2b68-406d-9d12-a690c3da4045-seedance_1.5_input.webp", "end_image_url": "https://segmind-resources.s3.amazonaws.com/output/dbc484a8-bcda-4b9c-97f6-d89dca2f3815-output-1767723467082.png", "duration": "5" } ``` ### Output Schema The API returns a synchronous response based on the model type: **For Image/Video/Audio Models**: - Response contains binary data (image/png, video/mp4, audio/mp3) - Content-Type header indicates the media type - Save the response body directly to a file **For Text Models**: - Response is JSON with the generated text - Structure varies by model **HTTP Response Codes**: - **200 - OK**: Request successful, output in response body - **400 - Bad Request**: Invalid parameters - **401 - Unauthorized**: Invalid or missing API key - **404 - Not Found**: Model not found - **406 - Not Acceptable**: Insufficient credits - **429 - Too Many Requests**: Rate limit exceeded - **500 - Server Error**: Internal server error ## About ### Kling Omni Video O1: Image-to-Video Model **Edited by Segmind Team on January 26, 2026.** --- #### What is Kling Omni Video O1? **Kling Omni Video O1**, by **Kuaishou**, is a cutting-edge image-to-video AI model that impressively brings still images to life by transforming them into cinematic animations while sticking to principles of physics for realistic output. It goes beyond conventional interpolation tools to effectively understand the nuances of scene dynamics to generate dynamic, realistic movements that appear authentic by 'respecting' the physics and spatial interactions. **Kling Omni Video O1** is built with the Multi-modal Visual Language (MVL) technology, hence it perfectly upholds subject consistency, ensuring character identity, props, color tones, and lighting across the entire video sequence. All these aspects make this model an asset to developers who can harness its creative power through a REST API to produce professional-quality videos at scale and even within a limited timeline. #### Key Features of Kling Omni Video O1 - **Subject Consistency Preservation**: It successfully maintains character identity, props, colors, and lighting across all frames. - **Physics-Based Motion**: It is capable of generating realistic animations that follow natural movement patterns. - **Dual-Image Control**: It accepts images for the start and (optional) end points for precise transformation guidance. - **Text-Guided Animation**: It accepts prompts to control motion direction, camera angles, and scene dynamics. - **Flexible Duration Options**: It offers the choice between 5-second (fast animations) or 10-second (slower, detailed) outputs. - **REST API Access**: It has a production-ready endpoint with no cold starts for consistent performance. - **Cinematic Output Quality**: It supports professional-grade video generation suitable for creative and commercial projects. #### Best Use Cases - **Content Creation & Marketing**: It is perfect to animate product shots, create dynamic social media content, or bring brand mascots to life while effectively maintaining a consistent visual identity. - **Film & Animation Pre-visualization**: It becomes an invaluable asset to quickly prototype scene transitions, character movements, or camera angles before committing to full production. - **E-commerce & Retail**: It will be a great tool to transform static product images into engaging visuals showing items in motion or use. - **Education & Training**: It is a powerful tool to convert diagrams and illustrations into animated demonstrations of processes or transformations. - **Game Development**: It is the go-to model to generate cutscene animations or concept videos from character art and environment designs. #### Prompt Tips and Output Quality - **Effective Prompt Structure**: Describe the transformation and the desired motion style to generate the best results. So, instead of "*car moves*," try the more descriptive prompt such as, "*sleek sports car accelerates forward with motion blur, camera tracking from side angle*." - **Image Quality Matters**: Use high-resolution start images with good lighting and clear subjects, as better source images produce sharper, more detailed videos while preserving input quality. - **End Image Usage**: Though it is optional, providing an end image gives the model a clear transformation target that further directs the results closer to your creative vision. This works particularly well for morphing effects or controlled state changes. - **Duration Selection**: Choose **5 seconds** for quick actions, product reveals, or high-energy animations; select **10 seconds** for gradual transformations, slower camera movements, or narrative-driven scenes. - **Motion Specificity**: It is always helpful to include camera angle descriptions ("dolly zoom," "pan left"), motion types ("gentle sway," "rapid spin"), and environmental effects ("wind blowing hair," "fabric rippling") for more controlled results. ## Usage Guide ### Kling Omni Video O1 Guide Unlock the full potential of Kling Omni Video O1 by following these best practices, prompt strategies, and parameter recommendations for various use cases. #### 1. Core Workflow 1. **Collect Your Assets** - Required: `start_image_url` (high-resolution, clear subject) - Optional: `end_image_url` (for precise morphs or state changes) 2. **Craft a Detailed Prompt** - Describe both **what** transforms and **how** it moves. - Include camera directions (“tracking shot,” “dolly zoom”), motion details (“gentle sway,” “rapid spin”), and environmental cues (“wind-blown hair,” “water rippling”). 3. **Choose Duration** - `5` seconds: Quick demos, high-energy reveals, or previews. - `10` seconds: Narrative scenes, slow pans, or elaborate transformations. #### 2. Prompt & Image Tips - **Be Specific** “Vintage motorcycle accelerates forward; camera tilts up as exhaust smoke drifts” beats “motorcycle moves.” - **Use Dual-Image Control** Add `end_image_url` to guide final state. Ideal for brand logo reveals morphing into product shots. - **Maintain Quality** Inputs ≥1024px on the longest side, good contrast, no heavy compression artifacts. The model preserves textures, color tones, and lighting. #### 3. Parameter Recommendations by Use Case | Use Case | Prompt Focus | Duration | End Image? | |--------------------------------|--------------------------------------|----------|------------| | Social Media Product Teaser | “Sleek bottle spins, label zooms in” | 5 | No | | Training Animation | “Diagram components assemble slowly” | 10 | Yes | | Game Cutscene Prototype | “Hero draws sword, camera pans left” | 10 | No | | E-commerce Demo | “Sneakers bounce, fabric ripples” | 5 | Yes | #### 4. Sample JSON Payload ```json { "prompt": "Antique clock hands rotate clockwise; camera dollies in slowly.", "start_image_url": "https://example.com/input.jpg", "end_image_url": "https://example.com/final.jpg", "duration": "10" } ``` #### 5. Performance Tips - Host images on a fast CDN to reduce upload latency. - For rapid iteration, start with a 5-second duration then scale up. - Monitor output and refine prompts—small changes in wording can dramatically affect motion style. By combining precise prompts, high-quality inputs, and appropriate duration settings, you’ll achieve cinematic, physics-aware animations that maintain visual consistency from frame to frame. Experiment with different camera angles, speeds, and dual-image scenarios to discover new creative possibilities! ## FAQ ### How does Kling Omni Video O1 differ from standard video interpolation? **Kling Omni Video O1** uses MVL technology to understand scene context and generate physics-based motion, thereby creating new frames with natural movement rather than morphing pixels; this aspect makes it significantly better than frame interpolation tools that simply blend between images. ### Can I control camera movement separately from subject motion? To achieve multi-layered control in camera movement, include camera directives in your prompt: e.g., "camera slowly zooms in," "tracking shot following subject", etc., alongside subject actions. ### What image formats and resolutions work best? **Kling Omni Video O1** accepts standard web image formats via URL. To achieve the high-end results, use images with clear subjects, good contrast, and resolution of at least 1024px on the longest side. ### Is an end image required for good results? No, the model works well with only a start image and text prompt. You may add an end image for precise transformation control or specific final states. ### How do I optimize for faster processing? Select the **5-second duration** option and ensure your start image URL is hosted on fast, reliable infrastructure to minimize upload time. ### What happens to small details in the original image? **Kling Omni Video O1's** subject consistency algorithms work efficiently to preserve fine details, textures, and color information throughout the animation; though extreme detail may soften slightly during rapid motion. ## Usage Examples ### cURL ```bash curl -X POST "https://api.segmind.com/v1/kling-o1-image-to-video" \ -H "x-api-key: YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "prompt": "A seamless transformation of a man driving a vehicle and crashing", "start_image_url": "https://segmind-resources.s3.amazonaws.com/output/d472b3e1-2b68-406d-9d12-a690c3da4045-seedance_1.5_input.webp", "end_image_url": "https://segmind-resources.s3.amazonaws.com/output/dbc484a8-bcda-4b9c-97f6-d89dca2f3815-output-1767723467082.png", "duration": "5" }' ``` ### Python ```python import requests import json api_key = "YOUR_API_KEY" url = "https://api.segmind.com/v1/kling-o1-image-to-video" data = { "prompt": "A seamless transformation of a man driving a vehicle and crashing", "start_image_url": "https://segmind-resources.s3.amazonaws.com/output/d472b3e1-2b68-406d-9d12-a690c3da4045-seedance_1.5_input.webp", "end_image_url": "https://segmind-resources.s3.amazonaws.com/output/dbc484a8-bcda-4b9c-97f6-d89dca2f3815-output-1767723467082.png", "duration": "5" } response = requests.post( url, json=data, headers={ 'x-api-key': api_key, 'Content-Type': 'application/json' } ) if response.status_code == 200: # For image/video/audio models, response.content contains the binary data with open('output.png', 'wb') as f: f.write(response.content) print('Generation complete, saved to output.png') else: print(f"Error: {response.status_code}") print(response.text) ``` ### JavaScript ```javascript const apiKey = 'YOUR_API_KEY'; const url = 'https://api.segmind.com/v1/kling-o1-image-to-video'; const data = { "prompt": "A seamless transformation of a man driving a vehicle and crashing", "start_image_url": "https://segmind-resources.s3.amazonaws.com/output/d472b3e1-2b68-406d-9d12-a690c3da4045-seedance_1.5_input.webp", "end_image_url": "https://segmind-resources.s3.amazonaws.com/output/dbc484a8-bcda-4b9c-97f6-d89dca2f3815-output-1767723467082.png", "duration": "5" }; const response = await fetch(url, { method: 'POST', headers: { 'x-api-key': apiKey, 'Content-Type': 'application/json', }, body: JSON.stringify(data), }); if (response.ok) { // For image/video/audio models, response contains binary data const blob = await response.blob(); const downloadUrl = URL.createObjectURL(blob); // Create download link const a = document.createElement('a'); a.href = downloadUrl; a.download = 'output.png'; a.click(); console.log('Generation complete'); } ``` ## Additional Resources ### Documentation - [Model Playground](https://www.segmind.com/models/kling-o1-image-to-video) - [API Documentation](https://www.segmind.com/models/kling-o1-image-to-video/api) - [Pricing Details](https://www.segmind.com/models/kling-o1-image-to-video/pricing) - [Platform Documentation](https://docs.segmind.com/)