# Google | Gemini Omni 1.1 Flash | Text to Video Gemini Omni 1.1 Flash Text-to-Video animates a source image into a short AI video with prompt-guided motion, synchronized audio, portrait or landscape framing, and selectable resolution up to 4k. ## API Information - **Model Slug:** google-gemini-omni-1-1-flash-text-to-video - **Branded URL:** https://www.eachlabs.ai/google/gemini-omni-flash/google-gemini-omni-1-1-flash-text-to-video - **Provider:** Google - **Category:** Text to Video - **Output Type:** video - **Status:** active - **Base Cost:** $1.50–$17.50 per million tokens, depending on token type - **Estimated Processing Time:** 75 seconds - **Interactive Demo:** https://www.eachlabs.ai/ai-models/google-gemini-omni-1-1-flash-text-to-video ## Pricing - **Charge Type:** dynamic - **Estimate:** $1.50–$17.50 per million tokens, depending on token type - **Pricing Details:** input tokens: $1.50 - **Pricing Details:** output tokens: $17.50 - **Pricing Details:** thought tokens: $9.00 ## Input Schema | Parameter | Type | Required | Default | Constraints | Description | |-----------|------|----------|---------|-------------|-------------| | prompt | string | Yes | - | - | Text description guiding the video. Describe the scene, motion, camera work, dialogue, sound effects, or ambient audio. Gemini Omni 1.1 Flash generates synchronized audio with video. Max 20000 characters. | | aspect_ratio | string | No | 16:9 | 16:9, 9:16 | Output video aspect ratio. 16:9 is landscape; 9:16 is portrait. | | resolution | string | No | 720p | 360p, 720p, 1080p, 4k | Output video resolution. Gemini Omni 1.1 Flash supports 360p, 720p, 1080p, and 4k. Defaults to 720p if omitted. | | duration | string | No | - | 3s, 4s, 5s, 6s, 7s, 8s, 9s, 10s | Optional output video length, 3-10 seconds. If omitted, the model chooses a length based on the prompt. | ## Example Request ```bash curl -X POST https://api.eachlabs.ai/v1/prediction/ \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "google-gemini-omni-1-1-flash-text-to-video", "input": { "prompt": "Ultra cinematic macro nature film, hyperrealistic, soft depth of field, slow motion, atmospheric, dreamlike, volumetric lighting, Fibonacci-inspired transitions, seamless morphing between natural spirals, elegant camera movement, documentary-quality realism" } }' ``` ## Output Schema Response returned by `GET /v1/prediction/{id}` when the job completes: ```json { "status": "success", "predictionID": "string", "output": "string (URL of generated video)", "metrics": { "predict_time": "number (seconds)" } } ``` ## Polling ```bash curl https://api.eachlabs.ai/v1/prediction/{PREDICTION_ID} \ -H "Authorization: Bearer YOUR_API_KEY" ``` | Status | Meaning | |--------|---------| | `processing` | Still running — poll again | | `success` | Done — read `output` | | `error` | Failed — read `message` / `details` | ## Webhook (alternative to polling) Pass `"webhook_url": "https://your.host/path"` in the create request. Eachlabs POSTs this payload when the job ends: ```json { "exec_id": "prediction-uuid", "status": "succeeded", "output": "https://...", "error": "" } ``` `status` is `"succeeded"` or `"failed"`. `exec_id` equals the `predictionID` from create. Return 2xx within 30 seconds. ## Errors Error body: `{ "status": "error", "message": "...", "details": "..." }` | Code | Meaning | |------|---------| | `400` | Invalid input | | `401` | Missing / invalid `Authorization` bearer token | | `404` | Unknown model or prediction id | | `429` | Rate limit — 100 creates / min, 10 concurrent per key | | `5xx` | Retry with backoff | ## Overview **Google | Gemini Omni 1.1 Flash | Text to Video Overview** Google | Gemini Omni 1.1 Flash | Text to Video is Google’s latest multimodal video generation model that turns text prompts and reference media into short, high-quality clips with synchronized audio. It is part of the Gemini Omni family from Google DeepMind, designed specifically for controllable video generation and editing. Compared with earlier Gemini Omni Flash releases, Gemini Omni 1.1 Flash adds support for higher resolutions up to 4K and more advanced scene extension controls, giving creators a faster way to prototype motion while keeping visual quality competitive with frontier text‑to‑video models. On each::labs, the Google | Gemini Omni 1.1 Flash | Text to Video model focuses on animating a single source image into short videos, guided by prompts that define motion, framing, and sound so users can move from static visuals to dynamic, share‑ready clips in a single workflow. ## Usage Notes - API Base URL: `https://api.eachlabs.ai/v1` - Authentication: send `Authorization: Bearer YOUR_API_KEY`. Generate a key from the Eachlabs dashboard at https://www.eachlabs.ai/dashboard/api-keys. - File-typed parameters (`*_url`, `image_url`, `video_url`, `audio_url`, etc.) accept publicly-reachable HTTPS URLs only. Upload your asset first (GCS / S3 / your CDN) and pass the resulting URL. Data-URIs and localhost URLs are rejected. - For structured parameters (arrays / objects) send real JSON values, not stringified payloads. - Monetary values are reported in USD; per-token / per-megapixel rates may be billed in micro-cents internally. - Prefer `webhook_url` over polling for long-running predictions — see the Webhook Callback section.