Pseudo Dataset Generation for Out-of-Domain Multi-Camera View Recommendation

Kuan-Ying Lee; Qian Zhou; Klara Nahrstedt

Pseudo Dataset Generation for Out-of-Domain Multi-Camera View Recommendation

Kuan-Ying Lee, Qian Zhou, Klara Nahrstedt

TL;DR

By training the model on pseudo-labeled datasets stemming from videos in the target domain, the model achieves a 68% relative improvement in the model’s accuracy in the target domain and bridge the accuracy gap between in-domain and never-before-seen domains.

Abstract

Multi-camera systems are indispensable in movies, TV shows, and other media. Selecting the appropriate camera at every timestamp has a decisive impact on production quality and audience preferences. Learning-based view recommendation frameworks can assist professionals in decision-making. However, they often struggle outside of their training domains. The scarcity of labeled multi-camera view recommendation datasets exacerbates the issue. Based on the insight that many videos are edited from the original multi-camera videos, we propose transforming regular videos into pseudo-labeled multi-camera view recommendation datasets. Promisingly, by training the model on pseudo-labeled datasets stemming from videos in the target domain, we achieve a 68% relative improvement in the model's accuracy in the target domain and bridge the accuracy gap between in-domain and never-before-seen domains.

Pseudo Dataset Generation for Out-of-Domain Multi-Camera View Recommendation

TL;DR

Abstract

Pseudo Dataset Generation for Out-of-Domain Multi-Camera View Recommendation

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (3)