Enhancing Rotation-Invariant 3D Learning with Global Pose Awareness and Attention Mechanisms

Jiaxun Guo; Manar Amayri; Nizar Bouguila; Xin Liu; Wentao Fan

Enhancing Rotation-Invariant 3D Learning with Global Pose Awareness and Attention Mechanisms

Jiaxun Guo, Manar Amayri, Nizar Bouguila, Xin Liu, Wentao Fan

TL;DR

This work addresses the limitation of rotation-invariant 3D learning where global pose cues are lost due to restricted receptive fields, causing indistinguishability of symmetric components (wing-tip collapse). It introduces Shadow-informed Pose Feature (SiPF), which augments local RI descriptors with a globally consistent shadow derived from a learnable rotation, and couples it with Rotation-invariant Attention Convolution (RIAttnConv) and a Bingham-distribution-based shadow locating module to preserve global pose information while maintaining RI. The approach yields state-of-the-art results on 3D classification and part segmentation under arbitrary rotations, demonstrating strong gains especially in fine-grained spatial discrimination and robustness to real-world noise. This method enhances practical 3D perception by enabling rotation-invariant processing that remains sensitive to global structure, with potential impact on autonomous systems and robotics where objects appear in unconstrained orientations.

Abstract

Recent advances in rotation-invariant (RI) learning for 3D point clouds typically replace raw coordinates with handcrafted RI features to ensure robustness under arbitrary rotations. However, these approaches often suffer from the loss of global pose information, making them incapable of distinguishing geometrically similar but spatially distinct structures. We identify that this limitation stems from the restricted receptive field in existing RI methods, leading to Wing-tip feature collapse, a failure to differentiate symmetric components (e.g., left and right airplane wings) due to indistinguishable local geometries. To overcome this challenge, we introduce the Shadow-informed Pose Feature (SiPF), which augments local RI descriptors with a globally consistent reference point (referred to as the 'shadow') derived from a learned shared rotation. This mechanism enables the model to preserve global pose awareness while maintaining rotation invariance. We further propose Rotation-invariant Attention Convolution (RIAttnConv), an attention-based operator that integrates SiPFs into the feature aggregation process, thereby enhancing the model's capacity to distinguish structurally similar components. Additionally, we design a task-adaptive shadow locating module based on the Bingham distribution over unit quaternions, which dynamically learns the optimal global rotation for constructing consistent shadows. Extensive experiments on 3D classification and part segmentation benchmarks demonstrate that our approach substantially outperforms existing RI methods, particularly in tasks requiring fine-grained spatial discrimination under arbitrary rotations.

Enhancing Rotation-Invariant 3D Learning with Global Pose Awareness and Attention Mechanisms

TL;DR

Abstract

Enhancing Rotation-Invariant 3D Learning with Global Pose Awareness and Attention Mechanisms

TL;DR

Abstract

Paper Structure

Table of Contents

Key Result

Figures (6)

Theorems & Definitions (16)