arXiv:2609.13151v1 Announce Type: new Abstract: Leading multilingual speech recognition models like Whisper transcribe diverse, low-resource languages without language-specific training but are computationally expensive to deploy. Token merging mitigates this inefficiency by…
Read the original source — arxiv.org
paper · Shared by tscosj
0 comments
No comments yet.