GenHPE: Generative Counterfactuals for 3D Human Pose Estimation with Radio Frequency Signals

Shuokang Huang; Julie A. McCann

GenHPE: Generative Counterfactuals for 3D Human Pose Estimation with Radio Frequency Signals

Shuokang Huang, Julie A. McCann

TL;DR

This work tackles cross-domain 3D human pose estimation from radio frequency signals by eliminating domain-specific confounders through generative counterfactuals. GenHPE synthesizes counterfactual RF signals conditioned on manipulated skeletons, computes differences to isolate body-part effects, and regularizes a domain-invariant encoder-decoder to improve generalization across unseen subjects and environments. The approach, validated on WiFi, UWB, and mmWave datasets, with DDPM, DDIM, and CGAN variants, delivers state-of-the-art cross-domain accuracy and shows ablations that confirm the importance of skeleton embeddings and counterfactual regularization. The method offers a practical, privacy-preserving alternative to camera-based HPE, with strong potential for robust sensing in diverse real-world environments.

Abstract

Human pose estimation (HPE) detects the positions of human body joints for various applications. Compared to using cameras, HPE using radio frequency (RF) signals is non-intrusive and more robust to adverse conditions, exploiting the signal variations caused by human interference. However, existing studies focus on single-domain HPE confined by domain-specific confounders, which cannot generalize to new domains and result in diminished HPE performance. Specifically, the signal variations caused by different human body parts are entangled, containing subject-specific confounders. RF signals are also intertwined with environmental noise, involving environment-specific confounders. In this paper, we propose GenHPE, a 3D HPE approach that generates counterfactual RF signals to eliminate domain-specific confounders. GenHPE trains generative models conditioned on human skeleton labels, learning how human body parts and confounders interfere with RF signals. We manipulate skeleton labels (i.e., removing body parts) as counterfactual conditions for generative models to synthesize counterfactual RF signals. The differences between counterfactual signals approximately eliminate domain-specific confounders and regularize an encoder-decoder model to learn domain-independent representations. Such representations help GenHPE generalize to new subjects/environments for cross-domain 3D HPE. We evaluate GenHPE on three public datasets from WiFi, ultra-wideband, and millimeter wave. Experimental results show that GenHPE outperforms state-of-the-art methods and reduces estimation errors by up to 52.2mm for cross-subject HPE and 10.6mm for cross-environment HPE.

GenHPE: Generative Counterfactuals for 3D Human Pose Estimation with Radio Frequency Signals

TL;DR

Abstract

GenHPE: Generative Counterfactuals for 3D Human Pose Estimation with Radio Frequency Signals

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (8)