Human Fall Detection Using Deep Learning Methods: A Multimodal Audio-Video Approach
This study proposes a multimodal deep learning framework that fuses audio and video features using decision-level stacking to achieve a 98% accuracy in human fall detection, significantly outperforming single-modality models and other fusion strategies.