StickMotion: Generating 3D Human Motions by Drawing a Stickman

Tao Wang; Zhihua Wu; Qiaozhi He; Jiaming Chu; Ling Qian; Yu Cheng; Junliang Xing; Jian Zhao; Lei Jin

StickMotion: Generating 3D Human Motions by Drawing a Stickman

Tao Wang, Zhihua Wu, Qiaozhi He, Jiaming Chu, Ling Qian, Yu Cheng, Junliang Xing, Jian Zhao, Lei Jin

TL;DR

This work tackles the challenge of generating 3D human motions from text by introducing StickMotion, a diffusion-based framework that jointly leverages textual descriptions and stickman cues placed at the start, middle, and end of a motion sequence. A Stickman Generation Algorithm automatically creates stickman representations, while a Multi-Condition Module fuses text and stickman inputs efficiently during diffusion, supported by a Dynamic Supervision strategy that aligns stickman positions with natural motion. The authors also propose the StiSim metric to quantify stickman influence and report competitive results on KIT-ML and HumanML3D, along with a user study showing about 51.5% time savings for sketch-based specification. Overall, StickMotion enables more intuitive, user-friendly control of 3D motion generation with reduced computational cost and validated effectiveness.

Abstract

Text-to-motion generation, which translates textual descriptions into human motions, has been challenging in accurately capturing detailed user-imagined motions from simple text inputs. This paper introduces StickMotion, an efficient diffusion-based network designed for multi-condition scenarios, which generates desired motions based on traditional text and our proposed stickman conditions for global and local control of these motions, respectively. We address the challenges introduced by the user-friendly stickman from three perspectives: 1) Data generation. We develop an algorithm to generate hand-drawn stickmen automatically across different dataset formats. 2) Multi-condition fusion. We propose a multi-condition module that integrates into the diffusion process and obtains outputs of all possible condition combinations, reducing computational complexity and enhancing StickMotion's performance compared to conventional approaches with the self-attention module. 3) Dynamic supervision. We empower StickMotion to make minor adjustments to the stickman's position within the output sequences, generating more natural movements through our proposed dynamic supervision strategy. Through quantitative experiments and user studies, sketching stickmen saves users about 51.5% of their time generating motions consistent with their imagination. Our codes, demos, and relevant data will be released to facilitate further research and validation within the scientific community.

StickMotion: Generating 3D Human Motions by Drawing a Stickman

TL;DR

Abstract

StickMotion: Generating 3D Human Motions by Drawing a Stickman

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (5)