Delayed-KD: Why Matching Every Frame Hurts Streaming Speech Recognition
A voice assistant that recognizes speech only after you finish talking is not very useful. One that recognizes speech as you talk, word by word, has to make each decision with less information than a model that gets to hear…
Delayed-KD: Why Matching Every Frame Hurts Streaming Speech Recognition Read More »

