Jelajahi dan integrasikan model AI gemini-omni-text-to-video melalui MuAPI. Dapatkan inferensi berkecepatan tinggi dan harga kompetitif.
About this model
Model gemini-omni-text-to-video menyediakan kemampuan generasi AI mutakhir di platform MuAPI.
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapi | $0.039–$0.39 per second of output (by resolution) | Billed per second of output video: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. Synchronized audio included at no extra charge. |
| Fal.ai | Harga per detik yang sebanding | 相同的底层模型(google/gemini-omni-flash/v1.1/text-to-video)。 |
| Replicate | Tidak tersedia | Gemini Omni is not currently available on Replicate. |
Billed per second of output video: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. Synchronized audio included at no extra charge.
相同的底层模型(google/gemini-omni-flash/v1.1/text-to-video)。
Gemini Omni is not currently available on Replicate.
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt Teks | string | 所需Video内容的文本描述。Gemini Omni Mendukung丰富的多模态Prompt teks,包括场景构图、镜头指令、对话和环境Audio提示。 | A clock begins ticking louder and rapidly grows larger, breaking through the table and floor. Its gears spin violently as the hands rotate uncontrollably. Walls crack apart as the giant clock expands until it fills the entire apartment. |
| Durasi (Detik)(秒) | Enum (4 options) | HasilkanVideo的Durasi(秒)。 | 8 |
| Resolusi | Enum (4 options) | OutputVideoResolusi。按Output秒数计费:360p 为 $0.039/秒,720p 为 $0.13/秒,1080p 为 $0.195/秒,4K 为 $0.39/秒。 | 1080p |
| Rasio Aspek | Enum (2 options) | OutputVideo的Rasio aspek。 | 16:9 |
| Audio / Suara ID | array | Up to 3 voice profile IDs returned by the Gemini Omni Audio endpoint. | - |
| 种子 | int | Nilai acak (seed)(0–2147483647)。固定后可复现结果,但由于模型随机性,结果仍可能有所不同。 | 0 |
| Karakter ID | array | Up to 3 character IDs from Gemini Omni Character to feature in the video. | - |
所需Video内容的文本描述。Gemini Omni Mendukung丰富的多模态Prompt teks,包括场景构图、镜头指令、对话和环境Audio提示。
A clock begins ticking louder and rapidly grows larger, breaking through the table and floor. Its gears spin violently as the hands rotate uncontrollably. Walls crack apart as the giant clock expands until it fills the entire apartment.HasilkanVideo的Durasi(秒)。
8OutputVideoResolusi。按Output秒数计费:360p 为 $0.039/秒,720p 为 $0.13/秒,1080p 为 $0.195/秒,4K 为 $0.39/秒。
1080pOutputVideo的Rasio aspek。
16:9Up to 3 voice profile IDs returned by the Gemini Omni Audio endpoint.
-Nilai acak (seed)(0–2147483647)。固定后可复现结果,但由于模型随机性,结果仍可能有所不同。
0Up to 3 character IDs from Gemini Omni Character to feature in the video.
-Developer documentation
Panduan integrasi cepat untuk gemini-omni-text-to-video. Hubungkan dengan kunci API MuAPI Anda dan mulai lakukan panggilan inferensi.
Frequently asked
Gemini Omni 是一个在一次前向传递中同时理解文本、图像、音频和视频的基础模型,而不是通过专用模型链转发输出。这样可以获得更干净的编辑、原生音频生成,以及更少的跨模态伪影。
会——同步对白、环境声和音乐会与画面在同一次处理中原生生成。如果想自行添加配乐,也可以禁用音频生成。
Gemini Omni 每次生成支持 4、6、8 或 10 秒。
共有四档:360p 用于快速、低成本草稿,720p 和 1080p 用于标准输出,4K 用于最终交付——每档都按输出秒数计费。
Google 于 2026 年 5 月 19 日的 I/O 2026 上宣布了 Gemini Omni,API 访问将在接下来几周陆续开放。API 上线后,muapi 将立即启用此端点——请关注更新页面的发布公告。