Skip to content

Compare models

Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.

  1. Whisper 1OpenAIRemove
whisper-1
AttributeWhisper 1whisper-1
Pricing
Input$75.00 / 1M
Output$75.00 / 1M
Cache Write (5m)Not applicable
Cache Write (1h)Not applicable
Cache ReadNot applicable
Web Search$0 / 1M
Context
Max contextN/A
Max outputN/A
Capabilities
VisionNo
Function CallingNo
JSON ModeYes
StreamingNo
Catalogue
ProviderOpenAI
Categoryvoice
Charge typePay As You Go
Released
Description
SummaryWhisper (whisper-1) is OpenAI's open-source automatic speech recognition (ASR) model, designed for audio transcription and translation. It supports 50+ languages and processes audio files up to 25 MB, accepting formats such as mp3, mp4, wav, and webm. Optimized for reliable speech-to-text conversion across diverse audio inputs, Whisper is priced per minute of audio, billed to the nearest second, making it well suited for transcription, localization, and voice-driven applications.