跳到正文
Apple Machine Learning Research·· 11 天前评分77

Compressing Streaming Neural Audio Encoders via Latent-Space Distillation

摘要

这条英文动态主要涉及模型能力与工程。原文要点:System-wide Dictation on Apple devices runs entirely on-device, and the speech it transcribes reaches the foundation model through a tokenizer: an encoder that maps short windows of waveform onto the representation the language model reads. Because that model is sparsely activated under Instruction-Following Pruning, only a small subset of its experts occupies DRAM at any time, so the always-on tokenizer competes for...

本站只提供摘要与原文入口。完整内容请阅读原文。

来源:Apple Machine Learning Research · machinelearning.apple.com