Whisper Tiny: Core ML encoder

This repository contains the optional Apple Silicon encoder for Whisper Tiny in Glimpse. It is a companion to a Whisper GGUF or whisper.cpp ggml model, not a replacement for it: the model file still provides the decoder, tokenizer and metadata, and this package moves the audio encoder onto Core ML. Glimpse uses it for every Whisper Tiny quantization.

File Download size Purpose
whisper-tiny-encoder.mlmodelc.zip 15.0 MB Compiled Core ML encoder

On Windows, Intel Macs, or anywhere without Core ML, use the model file on its own. The Core ML package requires Apple Silicon.

Provenance

The model originates from OpenAI Whisper Tiny, licensed Apache 2.0. The encoder was converted from whisper.cpp's F16 ggml-tiny.bin with scripts/convert-whisper-gguf-to-coreml.py from our transcribe.cpp fork, then zipped with ditto -c -k --keepParent.

Artifact SHA-256
Encoder ZIP 35041e7f9f9c3e016bf1ff23819109c9dbd60cb387f97526fd44ba26c916bf52

Runtime requirements

Extract the ZIP next to the model file, keeping the whisper-tiny-encoder.mlmodelc directory name. Glimpse-Speech looks for the companion beside the model and falls back to the model's own encoder when it isn't there.

The encoder follows Apple's Neural Engine transformer layout and computes in FP16. It takes one 30-second Whisper window. Core ML may run some operations on the CPU; this is not a guarantee of exclusive Neural Engine execution.

Glimpse

This encoder speeds up on-device dictation on Apple Silicon in Glimpse, a free, open-source dictation app for Mac and Windows. The source is on GitHub.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Glimpse-Dictation/Whisper-Tiny-coreml

Finetuned
(1918)
this model