3 dataset results for Motion Synthesis AND Texts

HumanML3D is a 3D human motion-language dataset that originates from a combination of HumanAct12 and Amass dataset. It covers a broad range of human actions such as daily activities (e.g., 'walking', 'jumping'), sports (e.g., 'swimming', 'playing golf'), acrobatics (e.g., 'cartwheel') and artistry (e.g., 'dancing'). Overall, HumanML3D dataset consists of 14,616 motions and 44,970 descriptions composed by 5,371 distinct words. The total length of motions amounts to 28.59 hours. The average motion length is 7.1 seconds, while average description length is 12 words.

119 PAPERS • 2 BENCHMARKS

KIT Motion-Language

The KIT Motion-Language is a dataset linking human motion and natural language.

35 PAPERS • 2 BENCHMARKS

InterHuman

InterHuman is a multimodal dataset, named InterHuman. It consists of about 107M frames for diverse two-person interactions, with accurate skeletal motions and 16,756 natural language descriptions.

15 PAPERS • 1 BENCHMARK

Datasets

3 dataset results for Motion Synthesis AND Texts