Robust Feature Learning for Multi-Index Models in High Dimensions

Alireza Mousavi-Hosseini; Adel Javanmard; Murat A. Erdogdu

Robust Feature Learning for Multi-Index Models in High Dimensions

Alireza Mousavi-Hosseini, Adel Javanmard, Murat A. Erdogdu

TL;DR

The paper addresses robust learning for high-dimensional data when the target depends on a low-dimensional projection (a multi-index model). It shows that under $\ell_2$ perturbations, the Bayes-optimal robust projection aligns with the standard low-dimensional subspace defined by the target directions $\boldsymbol{U}$, provided a mild independence assumption holds. A two-layer neural-network framework is proposed, where an oracle recovers $\boldsymbol{U}$ and a robust readout is trained on the projected, low-dimensional representation, yielding sample complexity that does not scale with the ambient dimension $d$. The work provides polynomial- and SFL/DFL-based guarantees, along with numerical experiments indicating practical benefits of pre-learning the latent subspace before adversarial tuning, and outlines open questions on extending to other perturbation models and deeper architectures.

Abstract

Recently, there have been numerous studies on feature learning with neural networks, specifically on learning single- and multi-index models where the target is a function of a low-dimensional projection of the input. Prior works have shown that in high dimensions, the majority of the compute and data resources are spent on recovering the low-dimensional projection; once this subspace is recovered, the remainder of the target can be learned independently of the ambient dimension. However, implications of feature learning in adversarial settings remain unexplored. In this work, we take the first steps towards understanding adversarially robust feature learning with neural networks. Specifically, we prove that the hidden directions of a multi-index model offer a Bayes optimal low-dimensional projection for robustness against $\ell_2$-bounded adversarial perturbations under the squared loss, assuming that the multi-index coordinates are statistically independent from the rest of the coordinates. Therefore, robust learning can be achieved by first performing standard feature learning, then robustly tuning a linear readout layer on top of the standard representations. In particular, we show that adversarially robust learning is just as easy as standard learning. Specifically, the additional number of samples needed to robustly learn multi-index models when compared to standard learning does not depend on dimensionality.

Robust Feature Learning for Multi-Index Models in High Dimensions

TL;DR

The paper addresses robust learning for high-dimensional data when the target depends on a low-dimensional projection (a multi-index model). It shows that under

perturbations, the Bayes-optimal robust projection aligns with the standard low-dimensional subspace defined by the target directions

, provided a mild independence assumption holds. A two-layer neural-network framework is proposed, where an oracle recovers

and a robust readout is trained on the projected, low-dimensional representation, yielding sample complexity that does not scale with the ambient dimension

. The work provides polynomial- and SFL/DFL-based guarantees, along with numerical experiments indicating practical benefits of pre-learning the latent subspace before adversarial tuning, and outlines open questions on extending to other perturbation models and deeper architectures.

Abstract

-bounded adversarial perturbations under the squared loss, assuming that the multi-index coordinates are statistically independent from the rest of the coordinates. Therefore, robust learning can be achieved by first performing standard feature learning, then robustly tuning a linear readout layer on top of the standard representations. In particular, we show that adversarially robust learning is just as easy as standard learning. Specifically, the additional number of samples needed to robustly learn multi-index models when compared to standard learning does not depend on dimensionality.

Robust Feature Learning for Multi-Index Models in High Dimensions

TL;DR

Abstract

Robust Feature Learning for Multi-Index Models in High Dimensions

TL;DR

Abstract

Paper Structure

Table of Contents

Key Result

Figures (1)

Theorems & Definitions (16)