使用 PixVerse V6 将任意图像制作成视频。支持最高 1080p 分辨率、最长 15 秒时长,以及基于الموجه النصي的运动控制。
About this model
PixVerse V6 صورة إلى فيديو可在文本الموجه النصي引导下,将任意静态图像制作成高质量视频。支持最高 1080p 分辨率、最长 15 秒时长、الموجه النصي优化模式以及可选的 AI 生成音频。
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | From $0.033/s (360p, no audio) to $0.150/s (1080p, with audio) | Per-second pricing. Default 720p without audio: $0.059/s. |
| Fal.ai | غير متوفر | PixVerse V6 image-to-video is not available on Fal.ai. |
| Replicate | غير متوفر | PixVerse V6 image-to-video is not available on Replicate. |
Per-second pricing. Default 720p without audio: $0.059/s.
PixVerse V6 image-to-video is not available on Fal.ai.
PixVerse V6 image-to-video is not available on Replicate.
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| الموجه النصي | string | 所需الفيديو运动和内容的文本描述。 | Cracks spread across the statue as it suddenly comes to life. Stone pieces fall off while glowing energy emerges from inside. The statue pulls itself free from the sand and takes a heavy step forward, shaking the ground as dust rises into the air. |
| الصورة URL | array | Upload or provide the input image to animate. | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/pixverse-v6-i2v.mp4 |
| الدقة | Enum (4 options) | المخرجاتالفيديو的الدقة。 | 720p |
| المدة (بالثواني)(秒) | int | الفيديوالمدة(秒)。 | 5 |
| الموجه النصي优化 | Enum (3 options) | 控制الموجه النصي增强。'enabled' 会重写الموجه النصي,'disabled' 按原样使用,'auto' 由模型决定。 | auto |
| 生成الصوت | boolean | Enable AI-generated audio for the video. | false |
所需الفيديو运动和内容的文本描述。
Cracks spread across the statue as it suddenly comes to life. Stone pieces fall off while glowing energy emerges from inside. The statue pulls itself free from the sand and takes a heavy step forward, shaking the ground as dust rises into the air.Upload or provide the input image to animate.
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/pixverse-v6-i2v.mp4المخرجاتالفيديو的الدقة。
720pالفيديوالمدة(秒)。
5控制الموجه النصي增强。'enabled' 会重写الموجه النصي,'disabled' 按原样使用,'auto' 由模型决定。
autoEnable AI-generated audio for the video.
falseDeveloper documentation
提供图像:通过 images_list 字段上传或链接单个图像 URL。
编写运动الموجه النصي:描述你希望出现的运动,例如镜头平移、主体移动或环境变化等。
选择分辨率和时长:选择所需的المخرجات质量(360p–1080p)和长度(1–15 秒)。
设置الموجه النصي优化:使用 thinking_type: auto 让模型决定是否增强الموجه النصي;使用 enabled 始终增强;使用 disabled 按原样使用الموجه النصي。
启用音频(可选):切换 generate_audio_switch,添加 AI 生成的环境声。
提交并轮询:API 返回 request_id。轮询 GET /api/v1/predictions/{request_id}/result,直到状态为 completed。
Frequently asked
必须通过 images_list 字段提供且只能提供一个图像 URL。
它控制الموجه النصي优化。'enabled' 会重写你的الموجه النصي以改善运动质量,'disabled' 会完全按照你写的الموجه النصي使用,而 'auto' 会让模型根据上下文自行决定。
费用按生成视频的秒数计算,并随分辨率变化。生成音频还会按秒收取附加费用。
支持任何可公开访问的图像 URL。常见格式包括 JPEG、PNG 和 WebP。