Controllable Dysarthric Speech Synthesis for Speaker-Diverse ASR Training
Researchers propose a speech synthesis method that separates speaker identity from dysarthric articulation patterns, allowing finer control over generated dysarthric speech. The approach conditions synthesis on individual patients, producing varied synthetic speakers to supplement scarce training data for dysarthric speech recognition. This addresses a field bottlenecked by high speaker variability and limited labeled recordings.