Whisper Base: Core ML encoder

This repository contains the optional Apple Silicon encoder for Whisper Base in Glimpse. It is a companion to a Whisper GGUF or whisper.cpp ggml model, not a replacement for it: the model file still provides the decoder, tokenizer and metadata, and this package moves the audio encoder onto Core ML. Glimpse uses it for every Whisper Base quantization.

File Download size Purpose
whisper-base-encoder.mlmodelc.zip 37.9 MB Compiled Core ML encoder

On Windows, Intel Macs, or anywhere without Core ML, use the model file on its own. The Core ML package requires Apple Silicon.

Provenance

The model originates from OpenAI Whisper Base, licensed Apache 2.0. The encoder was converted from whisper.cpp's F16 ggml-base.bin with scripts/convert-whisper-gguf-to-coreml.py from our transcribe.cpp fork, then zipped with ditto -c -k --keepParent.

Artifact SHA-256
Encoder ZIP f38ea79465a06476d59a7e60bba129df6fa0823264f22242c093b604f4c3e533

Runtime requirements

Extract the ZIP next to the model file, keeping the whisper-base-encoder.mlmodelc directory name. Glimpse-Speech looks for the companion beside the model and falls back to the model's own encoder when it isn't there.

The encoder follows Apple's Neural Engine transformer layout and computes in FP16. It takes one 30-second Whisper window. Core ML may run some operations on the CPU; this is not a guarantee of exclusive Neural Engine execution.

Glimpse

This encoder speeds up on-device dictation on Apple Silicon in Glimpse, a free, open-source dictation app for Mac and Windows. The source is on GitHub.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Glimpse-Dictation/Whisper-Base-coreml

Finetuned
(762)
this model