NVIDIA details a new workflow for evaluating clinical Automatic Speech Recognition (ASR) models faster, leveraging agent skills, NeMo Data Designer, and Nemotron Speech. This process enables the rapid, repeatable creation of pronunciation-aware synthetic audio for ASR benchmarks without requiring real patient data. It focuses on creating a repeatable feedback loop to improve clinical speech AI by addressing issues like rare terminology and pronunciation challenges.