# gemini-omni-image-to-video > Gemini Omni Image to Video — animate one or more reference images with a text prompt. Unified reasoning across modalities preserves subject identity and generates synchronized audio natively. ## Overview - **Endpoint**: `POST https://api.muapi.ai/api/v1/gemini-omni-image-to-video` - **Model ID**: `gemini-omni-image-to-video` - **Category**: image to video - **Variant**: Image to Video - **Family**: gemini-omni - **Cost**: 1.5 credits per call (some models compute cost dynamically based on params) ## API Usage MuApi uses a **submit-then-poll** pattern: submit a job, get a `request_id`, then poll the predictions endpoint until `status` is `completed`. Optionally pass `?webhook=YOUR_URL` on the submit call to receive a POST callback when the job finishes (skip polling). **Authentication**: send your MuApi key in the `x-api-key` header. Get one at https://muapi.ai/access-keys. ### 1. Submit a job ```http POST https://api.muapi.ai/api/v1/gemini-omni-image-to-video Content-Type: application/json x-api-key: YOUR_API_KEY ``` **Minimum (required only):** ```json { "prompt": "", "image_urls": "https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpg" } ``` **Full example (all params):** ```json { "prompt": "", "image_urls": "https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpg", "duration": 8, "resolution": "1080p", "aspect_ratio": "16:9", "audio_ids": [], "seed": 0, "character_ids": [] } ``` **Response:** ```json { "request_id": "abc123", "status": "processing" } ``` ### 2. Poll for the result ```http GET https://api.muapi.ai/api/v1/predictions/{request_id}/result x-api-key: YOUR_API_KEY ``` Possible `status` values: `queued`, `pending`, `processing`, `completed`, `failed`, `cancelled`. Poll every 2-5 seconds until terminal. When `completed`, the result URLs are in the `outputs` array. **Example response when `completed`:** ```json { "id": "abc123", "status": "completed", "outputs": [ "https://cdn.muapi.ai/.../output.png" ], "urls": { "get": "https://api.muapi.ai/api/v1/predictions/abc123/result" }, "created_at": "2026-05-08T12:34:56Z", "has_nsfw_contents": [] } ``` ### cURL ```bash # 1. Submit REQUEST_ID=$(curl -s -X POST https://api.muapi.ai/api/v1/gemini-omni-image-to-video \ -H "x-api-key: $MUAPI_API_KEY" \ -H "Content-Type: application/json" \ -d '{"prompt":"","image_urls":"https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpg"}' | jq -r .request_id) # 2. Poll until completed while :; do RESP=$(curl -s https://api.muapi.ai/api/v1/predictions/$REQUEST_ID/result -H "x-api-key: $MUAPI_API_KEY") STATUS=$(echo "$RESP" | jq -r .status) [ "$STATUS" = "completed" ] && echo "$RESP" | jq .outputs && break [ "$STATUS" = "failed" ] && echo "$RESP" && exit 1 sleep 3 done ``` ### Python ```python import os, time, requests API = "https://api.muapi.ai/api/v1" headers = {"x-api-key": os.environ["MUAPI_API_KEY"]} r = requests.post(f"{API}/gemini-omni-image-to-video", headers=headers, json={"prompt":"","image_urls":"https://cdn.muapi.ai/assets/gemini-omni-image-to-video.jpg"}) request_id = r.json()["request_id"] while True: res = requests.get(f"{API}/predictions/{request_id}/result", headers=headers).json() if res["status"] == "completed": print(res["outputs"]); break if res["status"] == "failed": raise RuntimeError(res.get("error")); time.sleep(3) ``` ## Input Schema The API accepts the following input parameters: - **`prompt`** (`string`, _required_): Text description of the desired motion and scene. Gemini Omni supports rich multimodal prompts including camera direction, dialogue, and ambient audio cues. - **`image_urls`** (`array`, _required_): Upload 1–7 reference images for the video. Maximum 20 MB each. - **`duration`** (`int`, _optional_): Duration of the generated video in seconds. - Default: `8` - Options: `"4"`, `"6"`, `"8"`, `"10"` - **`resolution`** (`string`, _optional_): Output video resolution. 720p and 1080p are the same price; 4K costs more. - Default: `"1080p"` - Options: `"720p"`, `"1080p"`, `"4k"` - **`aspect_ratio`** (`string`, _optional_): Output video aspect ratio. - Default: `"16:9"` - Options: `"16:9"`, `"9:16"` - **`audio_ids`** (`array`, _optional_): Up to 3 voice profile IDs returned by the Gemini Omni Audio endpoint. - **`seed`** (`int`, _optional_): Random seed (0–2147483647). Fix for reproducibility; results may still vary due to model stochasticity. - Default: `0` - **`character_ids`** (`array`, _optional_): Up to 3 character IDs from Gemini Omni Character to feature in the video. ## Output Schema The polling endpoint returns the following fields: - **`id`** (`string`): The request ID. - **`status`** (`string`): One of `queued`, `pending`, `processing`, `completed`, `failed`, `cancelled`. - **`outputs`** (`array`): URLs to generated images/videos/audio. Empty until `status` is `completed`. - **`urls.get`** (`string`): Self-link to re-fetch this prediction. - **`error`** (`string` | `null`): Error message if `status` is `failed`. - **`created_at`** (`string`): ISO-8601 timestamp of when the request was created. - **`has_nsfw_contents`** (`array of boolean`): Per-output NSFW detection flags. ## Webhooks (optional) Append `?webhook=https://your-server/path` to the submit URL. When the job reaches a terminal state, MuApi will POST the same shape as the polling response to your URL — no polling needed. ## Agent Integration MuApi ships an MCP server and CLI so agents (Claude Code, Cursor, custom) can call this endpoint without writing HTTP code: ```bash # Install the CLI npm install -g muapi-cli # Authenticate once muapi auth login # Expose all MuApi models as MCP tools to your agent muapi mcp serve ``` The MCP server exposes tools that wrap submit + poll for every model, including `gemini-omni-image-to-video`. See `muapi --help` for category-specific shortcuts (`muapi image generate`, `muapi video from-image`, etc.). ## Related Models - [Audio Profile](https://muapi.ai/playground/gemini-omni-audio) - [Image to Video](https://muapi.ai/playground/gemini-omni-image-to-video) - [Character Creation](https://muapi.ai/playground/gemini-omni-character) - [Text to Video](https://muapi.ai/playground/gemini-omni-text-to-video) - [Video Edit](https://muapi.ai/playground/gemini-omni-video-edit) ## Resources - [Playground Page](https://muapi.ai/playground/gemini-omni-image-to-video) - [API Reference](https://muapi.ai/playground/gemini-omni-image-to-video?tab=2) - [Global llms.txt](https://muapi.ai/llms.txt)