WildElder: A Chinese Elderly Speech Dataset from the Wild with Fine-Grained Manual Annotations
This paper introduces WildElder, a Mandarin elderly speech dataset collected from online videos with fine-grained manual annotations, designed to address the limitations of existing controlled-environment datasets and serve as a robust benchmark for automatic speech recognition and speaker profiling research.