Ovi 是统一的音频–视频生成模型,可将静态图像与描述性الموجه النصي转换为带同步音频的短视频。它同时支持نص إلى فيديو和图像条件视频المدخلات,并内置唇形同步、背景音频/音效及对白支持,让静态画面以电影感方式焕发生命力。视频以 540p 分辨率生成。
About this model
Ovi-image-to-video 是一款前沿 AI 模型,革新了将静态图像转换为动态视频内容的方式。通过整合先进的音视频合成技术,Ovi 可将静态图像与描述性الموجه النصي无缝结合,生成带同步声音、内置唇形同步和逼真背景音效的短小且引人入胜的视频。该创新模型同时支持نص إلى فيديو和图像条件المدخلات,是创意叙事与电影感内容创作的多用途工具。
Ovi 基于稳健的深度学习架构,凭借将静态视觉内容动画化并保持 540p 高质量المخرجات的独特能力脱颖而出。其技术实力不仅能驱动逼真的视听体验,还为希望提升多媒体影响力的企业和创作者提供易用且经济高效的解决方案。有了 Ovi,将想法转化为生动、动态的故事从未如此简单。
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.20 per generation | muapiapp offers the most cost-effective solution, being 20-50% cheaper than competitors while matching or exceeding quality standards. |
| Fal.ai | $0.30 per generation | Fal.ai charges slightly more, but muapiapp provides 20-50% savings with equally impressive performance and output quality. |
| Replicate | $0.30 per generation | Replicate’s pricing is on par with Fal.ai; however, muapiapp delivers the same high-quality video generation at a significantly lower cost. |
muapiapp offers the most cost-effective solution, being 20-50% cheaper than competitors while matching or exceeding quality standards.
Fal.ai charges slightly more, but muapiapp provides 20-50% savings with equally impressive performance and output quality.
Replicate’s pricing is on par with Fal.ai; however, muapiapp delivers the same high-quality video generation at a significantly lower cost.
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| الموجه النصي | string | 描述الفيديو内容的文本الموجه النصي。 | Camera: static medium shot. The scientist speaks: <S>We have discovered life beyond Earth.<E> <AUDCAP>Soft electronic hum, distant Beep of instruments<ENDAUDCAP> |
| الصورة URL | string | المدخلاتالصورة的 URL。 | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/ovi-image-to-video.jpg |
描述الفيديو内容的文本الموجه النصي。
Camera: static medium shot. The scientist speaks: <S>We have discovered life beyond Earth.<E> <AUDCAP>Soft electronic hum, distant Beep of instruments<ENDAUDCAP>المدخلاتالصورة的 URL。
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/ovi-image-to-video.jpgDeveloper documentation
准备المدخلات
المدخلات数据
生成视频
ovi-image-to-video endpoint。查看المخرجات
整合与分享
Frequently asked
Ovi-image-to-video 接受静态图像 URL 和描述性文本الموجه النصي。الموجه النصي可以包含详细的场景描述、带唇形同步提示的对白,以及用于增强最终视频المخرجات的音频说明。
该模型内置唇形同步和精确的音频对齐功能。它会分析描述性文本الموجه النصي,将对白和背景声音动态匹配到生成的视觉内容,从而提供连贯的视听体验。