ViPoser: Sparse-IMU Based Human Pose Estimation with Distilled Vision Foundation Priors
Published in ACM International Conference on Mobile Computing and Networking (MobiCom 2026), 2026
A lightweight framework that distills human-structure priors from a vision foundation model (Sapiens) into a compact IMU-based pose estimator — achieving 10–30% lower error and near-zero anatomically implausible poses under sparse 1–3 IMU input.
Recommended citation: Hanyu Zeng, et al. (2026). "ViPoser: Sparse-IMU Based Human Pose Estimation with Distilled Vision Foundation Priors." ACM MobiCom 2026.
Download Paper
