When ASR Models Learn the Test: New Methods Expose Benchmark Gaming
Researchers introduce three probes that catch speech models reproducing benchmark transcripts—even when the audio contradicts them, words are silenced, or spellings should vary randomly.