Whisper Base: Core ML encoder
This repository contains the optional Apple Silicon encoder for Whisper Base in Glimpse. It is a companion to a Whisper GGUF or whisper.cpp ggml model, not a replacement for it: the model file still provides the decoder, tokenizer and metadata, and this package moves the audio encoder onto Core ML. Glimpse uses it for every Whisper Base quantization.
| File | Download size | Purpose |
|---|---|---|
whisper-base-encoder.mlmodelc.zip |
37.9 MB | Compiled Core ML encoder |
On Windows, Intel Macs, or anywhere without Core ML, use the model file on its own. The Core ML package requires Apple Silicon.
Provenance
The model originates from OpenAI Whisper Base,
licensed Apache 2.0. The encoder was converted from whisper.cpp's F16 ggml-base.bin with
scripts/convert-whisper-gguf-to-coreml.py from our transcribe.cpp fork, then
zipped with ditto -c -k --keepParent.
| Artifact | SHA-256 |
|---|---|
| Encoder ZIP | f38ea79465a06476d59a7e60bba129df6fa0823264f22242c093b604f4c3e533 |
Runtime requirements
Extract the ZIP next to the model file, keeping the whisper-base-encoder.mlmodelc
directory name. Glimpse-Speech looks for the companion beside the model and falls back to the
model's own encoder when it isn't there.
The encoder follows Apple's Neural Engine transformer layout and computes in FP16. It takes one 30-second Whisper window. Core ML may run some operations on the CPU; this is not a guarantee of exclusive Neural Engine execution.
Glimpse
This encoder speeds up on-device dictation on Apple Silicon in Glimpse, a free, open-source dictation app for Mac and Windows. The source is on GitHub.
Model tree for Glimpse-Dictation/Whisper-Base-coreml
Base model
openai/whisper-base