Compare models
Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.
- Whisper 1OpenAIRemove
| Attribute | Whisper 1whisper-1 |
|---|---|
| Pricing | |
| Input | $75.00 / 1M |
| Output | $75.00 / 1M |
| Cache Write (5m) | Not applicable |
| Cache Write (1h) | Not applicable |
| Cache Read | Not applicable |
| Web Search | $0 / 1M |
| Context | |
| Max context | N/A |
| Max output | N/A |
| Capabilities | |
| Vision | No |
| Function Calling | No |
| JSON Mode | Yes |
| Streaming | No |
| Catalogue | |
| Provider | OpenAI |
| Category | voice |
| Charge type | Pay As You Go |
| Released | — |
| Description | |
| Summary | Whisper (whisper-1) is OpenAI's open-source automatic speech recognition (ASR) model, designed for audio transcription and translation. It supports 50+ languages and processes audio files up to 25 MB, accepting formats such as mp3, mp4, wav, and webm. Optimized for reliable speech-to-text conversion across diverse audio inputs, Whisper is priced per minute of audio, billed to the nearest second, making it well suited for transcription, localization, and voice-driven applications. |