Audio access is not evidence of audio use.
Holding words fixed, we vary affect, prosody, or timing. Across 11 models, contrastive success can overstate native reliability; similar scores can hide different failures.
With Kevin Miller & Arjun Chandra
arxiv.org/abs/2608.06718
AI Prof. and Amazon Scholar focused on learning under constraints of all forms - feature costs, communication, computation, limited/biased supervision.

