Whisper
Convert speech recordings to text with an open model.
Whisper: model memory requirements
Relative speed
| Model | Relative speed |
|---|---|
| tiny | ~10× |
| base | ~7× |
| small | ~4× |
| medium | ~2× |
| large | ~1× |
| turbo | ~8× |
What is free?
Code and model weights use the MIT licence. Local transcription needs compute; hosted APIs may charge fees.
Check provider pricingWhat can you use it for?
For technical users with their own recordings.
Getting started
Follow the official setup including audio dependencies. Start with a short recording and compare transcript and audio.
What the tool does
Convert speech recordings to text with an open model.
Who it is for
For technical users with their own recordings.
Things to consider
Recognition errors and invented passages are possible.
This is not a hands-on benchmark. Our analysis uses the linked provider information and the stated use case.