Every way
to use the major models.
Closed models like Claude and GPT — link to the cheapest API provider. Open-weights like Llama, Kimi, DeepSeek — choose hosted inference or self-host on rented GPUs.
5 models match — reset filters
Open-weights models.
Run yourself on cheap GPUs, or use a hosted-inference provider.
Whisper Large v3
OpenAI's open-weight speech-to-text — the standard transcription model.
Whisper Base
74M Whisper — browser / Raspberry Pi-deployable.
Whisper Medium
769M Whisper variant — half the size of Large, 80% of the accuracy.
Whisper Small
244M Whisper — fits on edge GPUs and CPU.
Whisper Tiny
39M Whisper — runs in-browser via WebGPU.