whisper.cpp
A dependency-free C/C++ port of OpenAI's Whisper speech recognition model that runs offline on CPUs, GPUs and mobile devices.
- GitHub stars
- 54k
- Last commit
- today
- Latest release
- v1.9.4
- Licence
- MIT
- Self-hosted
- Yes
whisper.cpp is a C/C++ implementation of OpenAI's Whisper automatic speech recognition model, built for efficient inference without external dependencies. It is released under the MIT license and part of the ggml project, and the whole high-level model implementation lives in two files, which makes it easy to embed in different platforms and applications.
It is tuned for many kinds of hardware: Apple Silicon through ARM NEON, Accelerate, Metal and Core ML, AVX on x86, VSX on POWER, plus Vulkan, NVIDIA GPUs, AMD ROCm, OpenVINO and several NPUs. Mixed F16 and F32 precision, integer quantization and zero memory allocations at runtime keep resource use low, and a C-style API and voice activity detection are included.
Supported platforms include macOS, iOS, Android, Java, Linux and FreeBSD, WebAssembly, Windows, Raspberry Pi and Docker. The README shows the model running fully offline on an iPhone and suggests building offline voice assistants. It suits developers adding private speech-to-text to apps, and hobbyists transcribing audio on their own hardware.
Key features
- Plain C/C++ implementation without dependencies
- Optimized for Apple Silicon, AVX and GPUs
- Integer quantization and mixed precision
- Zero memory allocations at runtime
- Voice activity detection
- C-style API for embedding
- Runs on mobile, desktop, web and Raspberry Pi
Pricing: Free and open source under the MIT license.
