ادغام مدل هوش مصنوعی gemini-omni-image-to-video از طریق API با کارایی بالا و تعرفه مصرف منعطف.
مدلها
دسترسی برنامهنویسی به مدل gemini-omni-image-to-video در پلتفرم MuAPI با کمترین تأخیر و مقیاسپذیری آنی.
تحلیل هزینه
| ارائهدهنده | هزینه | یادداشتها |
|---|---|---|
| muapi | $0.039–$0.39 به ازای هر ثانیه خروجی (by وضوح تصویر) | Billed به ازای هر ثانیه of output video: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. Synchronized audio included at no extra charge. |
| Fal.ai | تعرفه ثانیهای قابل مقایسه | Same underlying model (google/gemini-omni-flash/v1.1/image-to-video). |
| Replicate | در دسترس نیست | Gemini Omni Image to Video is در حال حاضر در دسترس نیست on Replicate. |
Billed به ازای هر ثانیه of output video: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. Synchronized audio included at no extra charge.
Same underlying model (google/gemini-omni-flash/v1.1/image-to-video).
Gemini Omni Image to Video is در حال حاضر در دسترس نیست on Replicate.
** مدلها。
طرح پیکربندی
| پارامتر | نوع | توضیحات | پیشفرض |
|---|---|---|---|
| پرامپت (دستور متنی) | string | تنظیم پارامتر Text description of the desired motion and scene. Gemini Omni supports rich multimodal prompts including camera direction, dialogue, and ambient audio cues. جهت هدایت دقیق فرآیند تولید مدل. | The suitcase opens by itself and tiny landscapes start unfolding out of it—mountains, forests, oceans, entire cities. Each world expands outward onto the platform, growing larger and larger while miniature weather systems form above them. |
| تصاویر مرجع | array | Upload 1–7 تصاویر مرجع for the video. Maximum 20 MB each. | https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpg |
| مدت زمان (ثانیه) | (4 ) | مدت زمان ویدیوی خروجی بر حسب ثانیه. | 8 |
| وضوح تصویر | (4 ) | تنظیم پارامتر Output video resolution. Billed per second of output: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. جهت هدایت دقیق فرآیند تولید مدل. | 1080p |
| نسبت ابعاد | (2 ) | نسبت ابعاد تصویر خروجی (مانند 16:9 یا 1:1). | 16:9 |
| شناسههای صوت | array | حداکثر تا 3 voice profile IDs returned by the Gemini Omni Audio endpoint. | - |
| مقدار سید تصادفی (Seed) | int | سید تصادفی (Seed) برای بازتولید دقیق خروجی. | 0 |
| شناسههای کاراکتر | array | حداکثر تا 3 character IDs from Gemini Omni Character to feature in the video. | - |
تنظیم پارامتر Text description of the desired motion and scene. Gemini Omni supports rich multimodal prompts including camera direction, dialogue, and ambient audio cues. جهت هدایت دقیق فرآیند تولید مدل.
The suitcase opens by itself and tiny landscapes start unfolding out of it—mountains, forests, oceans, entire cities. Each world expands outward onto the platform, growing larger and larger while miniature weather systems form above them.Upload 1–7 تصاویر مرجع for the video. Maximum 20 MB each.
https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpgمدت زمان ویدیوی خروجی بر حسب ثانیه.
8تنظیم پارامتر Output video resolution. Billed per second of output: $0.039/s at 360p, $0.13/s at 720p, $0.195/s at 1080p, $0.39/s at 4K. جهت هدایت دقیق فرآیند تولید مدل.
1080pنسبت ابعاد تصویر خروجی (مانند 16:9 یا 1:1).
16:9حداکثر تا 3 voice profile IDs returned by the Gemini Omni Audio endpoint.
-سید تصادفی (Seed) برای بازتولید دقیق خروجی.
0حداکثر تا 3 character IDs from Gemini Omni Character to feature in the video.
-مستندات توسعهدهندگان
یک درخواست POST به اندپوینت مدل همراه با کلید دسترسی MuAPI و پارامترهای مورد نظر ارسال فرمایید.
سوالات پرتکرار