Gemini Omni 视频编辑: AI Video Editor

Gemini Omni 视频编辑——原生多模态的视频到视频编辑。只需一个프롬프트,即可从源片段重新设计风格、调整光照、替换主体或改写场景。跨模态统一推理在应用编辑的同时保持运动和音频连续性。

📝

Overview

About this model

Gemini Omni 视频编辑将 Google 原生多模态 any-to-any 模型用于基于源片段的编辑。提供一个片段和自然语言编辑指令——重新设计外观、改变季节、替换主体或改写对白——模型就会在保留原始运动和时间安排的同时,通过一次处理重写视频。

由于模型会联合理解画面和音频,编辑可以在不同模态之间保持连贯:重新生成的音频会匹配新画面,需要时也可以保留原始环境声。

1重新设计风格与光照:通过一个프롬프트,将实拍片段转换为水彩、动漫、黏土动画或其他视觉风格。
2本地化:在保留说话者身份和唇部动作的同时,将对白改写为新语言。
3连续性修复:无需重新拍摄,即可调整现场遗漏的服装、道具或背景细节。
4创意迭代:从一个源片段测试同一镜头的多种外观——黄金时刻与霓虹夜景、夏季与冬季。
5内容适配:无需重新拍摄原始素材,即可为新平台重新构图和设计风格。
💰

Pricing & Value

Cost analysis

muapi$2.40 (720p/1080p) · $3.60 (4K)

Flat rate per generation — same price regardless of duration. Synchronized audio included at no extra charge.

Fal.ai제공되지 않음

Gemini Omni Video Edit is not currently available on Fal.ai.

Replicate제공되지 않음

Gemini Omni Video Edit is not currently available on Replicate.

** Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

编辑프롬프트string

描述要应用的编辑操作。例如:改为水彩风格、更换季节、替换主体服装。

Default ValueTransform the growing clock into @image1 while preserving the original expansion and wall-breaking motion. As the clock enlarges, metallic spider legs unfold from the sides and stab through furniture and walls. Keep the camera motion and apartment destruction intact.
参考이미지array

参考이미지。总共提供 7 个이미지槽位;비디오占用 2 个槽位,每个 character_id 占用 1 个槽位。每张이미지最大 20 MB。

Default Valuehttps://cdn.muapi.ai/assets/gemini-omni-video-edit.jpg
源비디오string

Source video to edit (max 100 MB, max 30 s). Optional if image_urls are provided.

Default Valuehttps://cdn.muapi.ai/assets/gemini-omni-text-to-video.mp4
裁剪开始时间(秒)number

Start of the clip window to edit, in seconds (>= 0).

Default Value0
裁剪结束时间(秒)number

End of the clip window to edit, in seconds. Must be within 10 s of trim_start.

Default Value10
재생 시간 (초)(秒)Enum (4 options)

生成비디오的재생 시간(秒)。

Default Value8
해상도Enum (3 options)

출력비디오해상도。720p 和 1080p 价格相同;4K 价格更高。

Default Value1080p
화면 비율Enum (2 options)

출력비디오的화면 비율。

Default Value16:9
오디오 IDarray

Up to 3 voice profile IDs returned by the Gemini Omni Audio endpoint.

Default Value-
种子int

랜덤 시드(0–2147483647)。固定后可复现结果,但由于模型随机性,结果仍可能有所不同。

Default Value0
角色 IDarray

Up to 3 character IDs from Gemini Omni Character to feature in the video.

Default Value-
📖

Implementation Guide

Developer documentation

사용 방법 Gemini Omni 视频编辑

  1. 上传源视频 提供公开的 video_url,或从 playground 上传片段。较短的입력 파라미터(约少于 15 秒)能产生最稳定的编辑结果。

  2. 编写清晰的编辑프롬프트 明确说明哪些内容应该改变、哪些内容应该保留。例如:'Restyle the entire clip as a hand-drawn Studio Ghibli animation, keep the original camera motion, lighting direction, and timing.'

  3. 选择分辨率

    • 720p / 1080p — 价格相同(每次生成 $2.40)
    • 4K — 更高分辨率(每次生成 $3.60)
  4. 提交并轮询 POST 到 /api/v1/gemini-omni-video-edit,并轮询 GET /api/v1/predictions/{request_id}/result,直到 statuscompleted

  5. 프롬프트技巧

    • 锚定不应改变的内容:'keep the original framing'、'preserve subject identity'、'maintain camera path'
    • 对于风格迁移,明确写出参考风格:'1980s VHS'、'oil painting'、'claymation'
    • 对于对白编辑,引用替换后的台词,并说明所需的声音语气或语言。

Common Questions

Frequently asked

这与将源片段作为参考运行텍스트 투 비디오模型有何不同?

Gemini Omni 视频编辑会直接以源片段为条件,同时理解画面和音频。应用编辑时,模型会保留源片段的运动、时间安排和连续性,而不是生成一个大致匹配的新片段。

可以只编辑音频(例如更改对白)而不改变画面吗?

可以——禁用 `preserve_audio`,并编写只针对音频的프롬프트(例如,"replace the dialogue with: '...'")。画面会与源片段保持对齐,同时重新生成音频。

支持多长的视频?

最大片段长度将在上线时确认。预计支持最长 30 秒的입력 파라미터,与텍스트 투 비디오版本一致。

编辑会影响运动或镜头移动吗?

默认情况下,模型会保留源片段的运动和镜头路径。如果希望改变镜头或运动,请在프롬프트中明确说明。

Gemini Omni 视频编辑何时会在 muapi 上线?

Google 于 2026 年 5 月 19 日的 I/O 2026 上宣布了 Gemini Omni,API 访问将在接下来几周陆续开放。视频编辑版本将在上游 API 支持后立即在 muapi 上线。