| Shen, Xiaojian, Shi, Dahu, Zhang, Jianrong, Li, Hai, Zhao, Hongwei, Zhang, Dawei, Zhuge, Yunzhi, Wu, Zhiliang, Yue, Guanghui and Zhou, Wei 2026. BeatDance: Generating beat-consistent 3D dance with hierarchical spatial–temporal modeling. Pattern Recognition 180 , 114344. 10.1016/j.patcog.2026.114344 |
Abstract
Generating realistic 3D dance from music is a challenging task that requires accurate synchronization with musical rhythms while capturing the spatial complexity of human motion. Although existing methods can generate physically plausible dance motions, they often struggle to achieve precise alignment with music, such as the beat. To address this limitation, we propose a novel diffusion-based framework, BeatDance, with two components: (1) We present a Hierarchical Decoupled Attention (HDA) module, which first disentangles the learning of human pose and temporal dynamics. A hierarchical structure is then employed to capture both short-term and long-term dependencies, thereby enhancing spatial–temporal modeling. (2) We adopt cycle-consistent learning by introducing an auxiliary dance-to-music module. During training, discrepancies between the reconstructed and original music induce a stronger loss signal, effectively encouraging the consistency property between the music and dance motion. Extensive experimental results demonstrate that our proposed approach outperforms recent competitive methods on two benchmark datasets. The project page is available at https://shenxiaojian.github.io/BeatDance/
| Item Type: | Article |
|---|---|
| Date Type: | Publication |
| Status: | Published |
| Schools: | Schools > Computer Science & Informatics |
| Publisher: | Elsevier |
| ISSN: | 0031-3203 |
| Date of Acceptance: | 22 June 2026 |
| Last Modified: | 09 Jul 2026 14:45 |
| URI: | https://orca.cardiff.ac.uk/id/eprint/188115 |
Actions (repository staff only)
![]() |
Edit Item |




Dimensions
Dimensions