Compressing Streaming Neural Audio Encoders via Latent-Space Distillation

Apple Machine Learning Research Version 1 original current

Imported from official source

System-wide Dictation on Apple devices runs entirely on-device, and the speech it transcribes reaches the foundation model through a tokenizer: an encoder that maps short windows of waveform onto the representation the language model reads. …

This version

Version
1 of 1
Recorded
September 24, 2026 16:00
Change
Initial
Content hash
60ab817c22132833663137dc74410b8e
All versions
Revision history

Officially records where a publication came from, not whether it is true. Imported records are reproduced from an organization's own official source.