Kling O1’s Reference-to-Video mode generates a dynamic video using one or multiple reference images as the visual foundation. It preserves identity, style, composition, and key visual details from the references while adding realistic camera motion, environment dynamics, and scene animation.
About this model
Kling O1’s Reference-to-Video model is a cutting-edge tool that transforms reference images into dynamic videos with remarkable precision. By preserving the identity, style, composition, and key visual details of the original images, the model seamlessly integrates realistic camera movement and environmental dynamics to create engaging visual narratives. This state-of-the-art technology leverages advanced AI algorithms and animation techniques to ensure that every generation feels authentic and immersive.
Built specifically for the Image to Video category, Kling O1 offers users the ability to control various parameters such as aspect ratio, video duration, and whether to maintain original sound. With an easy-to-use technical input schema and a competitive cost of just $0.72 per generation, this model stands out as a reliable and efficient solution for creative professionals and marketers seeking to bring their visual concepts to life with superior detail and dynamic motion.
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.72 | muapiapp offers this service at $0.72 per generation, making it 20-50% more affordable than its competitors, without compromising on quality. |
| Fal.ai | $0.95 | Fal.ai charges $0.95 per generation. Compared to Fal.ai, muapiapp is approximately 24% cheaper while delivering comparable or superior output quality. |
| Replicate | $0.95 | Replicate also offers similar functionalities at $0.95 per generation. In this competitive landscape, muapiapp provides a more cost-effective solution, being around 24% less expensive. |
muapiapp offers this service at $0.72 per generation, making it 20-50% more affordable than its competitors, without compromising on quality.
Fal.ai charges $0.95 per generation. Compared to Fal.ai, muapiapp is approximately 24% cheaper while delivering comparable or superior output quality.
Replicate also offers similar functionalities at $0.95 per generation. In this competitive landscape, muapiapp provides a more cost-effective solution, being around 24% less expensive.
* Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt | string | The prompt to generate the video | Cinematic orbit camera move around the pilot in a futuristic hangar, holographic lights flickering, armor reflections shifting, soft mechanical ambience. |
| Image URLs | array | Upload or provide image urls. Used for image-to-video generation. | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/kling-o1-reference-to-video-1.jpg |
| Video URL | string | URL of the input video. | null |
| Aspect Ratio | Enum (3 options) | Aspect ratio of the output video. | 16:9 |
| Duration | int | The duration of the generated video in seconds | 5 |
| Keep Original Sound | boolean | Select whether to keep the video original sound through the parameter. | true |
The prompt to generate the video
Cinematic orbit camera move around the pilot in a futuristic hangar, holographic lights flickering, armor reflections shifting, soft mechanical ambience.Upload or provide image urls. Used for image-to-video generation.
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/kling-o1-reference-to-video-1.jpgURL of the input video.
nullAspect ratio of the output video.
16:9The duration of the generated video in seconds
5Select whether to keep the video original sound through the parameter.
trueDeveloper documentation
How to Use Kling O1’s Reference-to-Video Mode
Prepare Your Inputs:
Set Up Your Parameters:
16:9, 9:16, or 1:1), duration (between 3 to 10 seconds), and decide if you want to keep the original sound.Generate the Video:
Review and Iterate:
Frequently asked
The model excels in preserving the identity, style, and key visual details of the input images while adding complex camera movements and scene animations. This ensures that each generated video is both dynamic and faithful to the original visuals.
Users can select from preset aspect ratios (`16:9`, `9:16`, `1:1`) to match their specific needs, ensuring the output video fits various platforms and formats.
Yes, there's a boolean parameter 'keep_original_sound' that allows you to decide if you want the video to retain its original audio track.
Each video generation using this model costs $0.72, making it a cost-effective solution for high-quality image-to-video transformations.