SoundMHPE Framework for Sound-Based Multi-Person 3D Pose Estimation
September 7, 2026
The SoundMHPE encoder-decoder framework enables 3D human pose estimation of multiple individuals using only acoustic signals. The architecture uses a multi-scale encoder to separate overlapping acoustic signatures and inter-person reflections in complex environments.
HOW THIS AFFECTS YOU
●
builderThis could enable non-visual human tracking in environments where cameras are restricted or privacy-sensitive.
●
researcherThis introduces a new modality for multi-subject pose estimation research using acoustic signal superposition.