使用 Kling 3 Omni Standard 从文本生成 720p 视频。在الموجه النصي中使用 <<<image_N>>> 添加最多 4 张参考图像,进行多图像引导的视频生成。免费试用——按生成次数付费,无需订阅。
About this model
Kling v3 Omni Standard 720P(T2V)。多图像参考视频生成——在الموجه النصي中使用 <<<image_N>>>,最多引用 4 张图像。通过 apimart 管理的 Kling 端点路由。
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.084/sec base / $0.112/sec with audio | 默认 5 秒片段费用为 $0.4200,与 Kling 官方定价一致。 |
| Kling.ai (official) | نفس السعر | muapiapp 通过 apimart 的 Kling 服务提供模型,价格与 Kling.ai 公布的费率一致。 |
| Replicate / Fal.ai | غير متوفر | Replicate 或 Fal.ai 尚未普遍提供 Kling v3 Omni。 |
默认 5 秒片段费用为 $0.4200,与 Kling 官方定价一致。
muapiapp 通过 apimart 的 Kling 服务提供模型,价格与 Kling.ai 公布的费率一致。
Replicate 或 Fal.ai 尚未普遍提供 Kling v3 Omni。
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| الموجه النصي | string | 文本الموجه النصي。通过 <<<image_N>>>(从 1 开始的索引)引用参考الصورة。如果省略,系统会自动在开头添加 <<<image_1>>>。 | A cyberpunk samurai crouched on a rooftop edge during a thunderstorm, glowing katana in hand. The samurai instantly launches forward across rooftops at extreme speed. Rain sprays behind each landing while he slices through neon signs and wall-runs across skyscrapers. The camera whips aggressively around every movement. |
| نسبة العرض إلى الارتفاع | Enum (3 options) | المخرجاتالفيديو的نسبة العرض إلى الارتفاع。 | 16:9 |
| المدة (بالثواني) | Enum (13 options) | 生成الفيديو的المدة(秒)。 | 5 |
| 生成الصوت | boolean | 启用后会随الفيديو生成原生الصوت(会增加费用)。 | false |
文本الموجه النصي。通过 <<<image_N>>>(从 1 开始的索引)引用参考الصورة。如果省略,系统会自动在开头添加 <<<image_1>>>。
A cyberpunk samurai crouched on a rooftop edge during a thunderstorm, glowing katana in hand. The samurai instantly launches forward across rooftops at extreme speed. Rain sprays behind each landing while he slices through neon signs and wall-runs across skyscrapers. The camera whips aggressively around every movement.المخرجاتالفيديو的نسبة العرض إلى الارتفاع。
16:9生成الفيديو的المدة(秒)。
5启用后会随الفيديو生成原生الصوت(会增加费用)。
falseDeveloper documentation
编写الموجه النصي:清晰描述场景。
选择نسبة العرض إلى الارتفاع和时长:16:9 / 9:16 / 1:1,3–15 秒。
切换 generate_audio:默认禁用。启用后会提高每秒费率。
提交并轮询:你会立即收到 request_id。轮询结果端点,直到状态为 completed。
Frequently asked
使用 `<<<image_1>>>`、`<<<image_2>>>` 等标记——它们按你在 `images_list` 中提交的顺序从 1 开始编号。如果跳过引用,模型会自动在الموجه النصي前添加 `<<<image_1>>>`。
需要——基础费率为 $0.084/sec;启用音频后为 $0.112/sec。
最长为 15 秒,最短为 3 秒。