# Gemini 2.5 Flash Gemini 2.5 is a fast and lightweight multimodal AI model designed for high performance and low latency, supporting text, image, audio, video, and document inputs while delivering quick responses with minimal resource usage. It is optimized for high-volume tasks such as classification, data extraction, and real-time applications, offering strong cost efficiency and scalable performance. ## API Information - **Model Slug:** gemini-2-5-flash - **Branded URL:** https://www.eachlabs.ai/google/gemini/gemini-2-5-flash - **Provider:** Google - **Category:** Text to Text - **Output Type:** text - **Status:** active - **Version:** 0.0.1 - **Base Cost:** Text/image/video input at $0.30/1M, audio input at $1.00/1M, output (incl. thinking) at $2.50/1M (Standard Paid tier) - **Estimated Processing Time:** 1 seconds - **Last Updated:** 2026-07-13 - **Interactive Demo:** https://www.eachlabs.ai/ai-models/gemini-2-5-flash ## Pricing - **Charge Type:** dynamic - **Pricing Details:** Text/image/video input at $0.30/1M, audio input at $1.00/1M, output (incl. thinking) at $2.50/1M (Standard Paid tier) ### Pricing Rules | Condition | Pricing | | --- | --- | | Rule 1 | Text/image/video input at $0.30/1M, audio input at $1.00/1M, output (incl. thinking) at $2.50/1M (Standard Paid tier) | ## Input Schema No input parameters documented. ## Example Request ```bash curl -X POST https://api.eachlabs.ai/v1/prediction/ \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "gemini-2-5-flash", "input": {} }' ``` ## Output Schema Response returned by `GET /v1/prediction/{id}` when the job completes: ```json { "status": "success", "predictionID": "string", "output": "text", "metrics": { "predict_time": "number (seconds)" } } ``` ## Polling ```bash curl https://api.eachlabs.ai/v1/prediction/{PREDICTION_ID} \ -H "Authorization: Bearer YOUR_API_KEY" ``` | Status | Meaning | |--------|---------| | `processing` | Still running — poll again | | `success` | Done — read `output` | | `error` | Failed — read `message` / `details` | ## Webhook (alternative to polling) Pass `"webhook_url": "https://your.host/path"` in the create request. Eachlabs POSTs this payload when the job ends: ```json { "exec_id": "prediction-uuid", "status": "succeeded", "output": "https://...", "error": "" } ``` `status` is `"succeeded"` or `"failed"`. `exec_id` equals the `predictionID` from create. Return 2xx within 30 seconds. ## Errors Error body: `{ "status": "error", "message": "...", "details": "..." }` | Code | Meaning | |------|---------| | `400` | Invalid input | | `401` | Missing / invalid `Authorization` bearer token | | `404` | Unknown model or prediction id | | `429` | Rate limit — 100 creates / min, 10 concurrent per key | | `5xx` | Retry with backoff | ## Overview **Gemini 2.5 Flash Overview** Google's **Gemini 2.5 Flash** is a fast, lightweight multimodal AI model from the Gemini family, optimized for high-volume text-to-text tasks with support for image, audio, video, and document inputs. It delivers quick responses with low latency, making it ideal for real-time applications like classification, data extraction, and content generation. As part of the Gemini 2.0 series, which emphasizes agentic AI capabilities, **Gemini 2.5 Flash** stands out for its efficiency in processing diverse inputs while maintaining strong performance in productivity workflows. Developed by Google, this model powers integrations in tools like Gmail, Docs, and Drive, enabling tasks such as summarizing documents or analyzing media. Its primary differentiator is the balance of speed and multimodality, allowing developers to build scalable **Gemini 2.5 Flash API** applications for **Google text-to-text** needs without heavy resource demands. Available via the Gemini API, it supports over 70 languages and operates in 230+ countries, positioning it as a cost-effective choice for high-throughput scenarios. ## Usage Notes - API Base URL: `https://api.eachlabs.ai/v1` - Authentication: send `Authorization: Bearer YOUR_API_KEY`. Generate a key from the Eachlabs dashboard at https://www.eachlabs.ai/dashboard/api-keys. - File-typed parameters (`*_url`, `image_url`, `video_url`, `audio_url`, etc.) accept publicly-reachable HTTPS URLs only. Upload your asset first (GCS / S3 / your CDN) and pass the resulting URL. Data-URIs and localhost URLs are rejected. - For structured parameters (arrays / objects) send real JSON values, not stringified payloads. - Monetary values are reported in USD; per-token / per-megapixel rates may be billed in micro-cents internally. - Prefer `webhook_url` over polling for long-running predictions — see the Webhook Callback section.