# Wizper Wizper is a multilingual speech recognition and translation model based on Whisper v3 that quickly and accurately converts audio files into text. It is optimized for real-time transcription and translation. ## API Information - **Model Slug:** wizper - **Branded URL:** https://www.eachlabs.ai/openai/whisper/wizper - **Provider:** Openai - **Category:** Voice to Text - **Output Type:** text - **Status:** active - **Version:** 0.0.1 - **Base Cost:** $0.00108 per processing second - **Estimated Processing Time:** 1 seconds - **Last Updated:** 2026-07-06 - **Interactive Demo:** https://www.eachlabs.ai/ai-models/wizper ## Pricing - **Charge Type:** dynamic - **Estimate:** $0.00108 per processing second - **Pricing Details:** default: $0.00108 ## Input Schema | Parameter | Type | Required | Default | Constraints | Description | |-----------|------|----------|---------|-------------|-------------| | audio_url | string | Yes | - | - | URL of the audio file to transcribe. Supported formats: mp3, mp4, mpeg, mpga, m4a, wav or webm. | | task | string | No | transcribe | transcribe,translate | Task to perform on the audio file. Either transcribe or translate. | | language | string | No | - | af,am,ar,as,az,ba,be,bg,bn,bo,br,bs,ca,cs,cy,da,de,el,en,es,et,eu,fa,fi,fo,fr,gl,gu,ha,haw,he,hi,hr,ht,hu,hy,id,is,it,ja,jw,ka,kk,km,kn,ko,la,lb,ln,lo,lt,lv,mg,mi,mk,ml,mn,mr,ms,mt,my,ne,nl,nn,no,oc,pa,pl,ps,pt,ro,ru,sa,sd,si,sk,sl,sn,so,sq,sr,su,sv,sw,ta,te,tg,th,tk,tl,tr,tt,uk,ur,uz,vi,yi,yo,zh | Language of the audio file. If translate is selected as the task, the audio will be translated to English, regardless of the language selected. If None is passed, the language will be automatically detected. This will also increase the inference time. Default value: en | | chunk_level | string | No | segment | - | Level of the chunks to return. | | version | string | No | 3 | - | Version of the model to use. All of the models are the Whisper large variant. | ## Example Request ```bash curl -X POST https://api.eachlabs.ai/v1/prediction/ \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "wizper", "input": { "audio_url": "https://cdn-us.eachlabs.ai/defaults/a8c3c8b8f7f893180f019703d657b7e5dec788d068ebe320b1eacc8db006d7ee.wav" } }' ``` ## Output Schema Response returned by `GET /v1/prediction/{id}` when the job completes: ```json { "status": "success", "predictionID": "string", "output": "text", "metrics": { "predict_time": "number (seconds)" } } ``` ## Polling ```bash curl https://api.eachlabs.ai/v1/prediction/{PREDICTION_ID} \ -H "Authorization: Bearer YOUR_API_KEY" ``` | Status | Meaning | |--------|---------| | `processing` | Still running — poll again | | `success` | Done — read `output` | | `error` | Failed — read `message` / `details` | ## Webhook (alternative to polling) Pass `"webhook_url": "https://your.host/path"` in the create request. Eachlabs POSTs this payload when the job ends: ```json { "exec_id": "prediction-uuid", "status": "succeeded", "output": "https://...", "error": "" } ``` `status` is `"succeeded"` or `"failed"`. `exec_id` equals the `predictionID` from create. Return 2xx within 30 seconds. ## Errors Error body: `{ "status": "error", "message": "...", "details": "..." }` | Code | Meaning | |------|---------| | `400` | Invalid input | | `401` | Missing / invalid `Authorization` bearer token | | `404` | Unknown model or prediction id | | `429` | Rate limit — 100 creates / min, 10 concurrent per key | | `5xx` | Retry with backoff | ## Overview **wizper — Voice-to-Text AI Model** Wizper, an advanced iteration in OpenAI's **Whisper** family based on Whisper v3, delivers fast and precise multilingual speech recognition, converting audio files to text with support for real-time transcription and translation. Developers and creators seeking a **voice-to-text AI model** rely on wizper to handle diverse accents and languages accurately, minimizing errors in noisy environments or live streams. Optimized for efficiency, it processes mono audio at 16 kHz sample rates, enabling seamless integration into apps for **OpenAI voice-to-text** workflows without quality loss. ## Usage Notes - API Base URL: `https://api.eachlabs.ai/v1` - Authentication: send `Authorization: Bearer YOUR_API_KEY`. Generate a key from the Eachlabs dashboard at https://www.eachlabs.ai/dashboard/api-keys. - File-typed parameters (`*_url`, `image_url`, `video_url`, `audio_url`, etc.) accept publicly-reachable HTTPS URLs only. Upload your asset first (GCS / S3 / your CDN) and pass the resulting URL. Data-URIs and localhost URLs are rejected. - For structured parameters (arrays / objects) send real JSON values, not stringified payloads. - Monetary values are reported in USD; per-token / per-megapixel rates may be billed in micro-cents internally. - Prefer `webhook_url` over polling for long-running predictions — see the Webhook Callback section.