arxivcs.CVcs.AI2026-07-31
MoRAE: Flow-Friendly Self-Supervised Latents for Text-to-Motion Generation
Yifei Zhu, Mingyi Shi, Yangyang Cai, Miao Cheng, Yoshifumi Kitamura, Taku Komura
Text-to-motion generation must produce motions that are semantically correct, temporally coherent, and physically plausible. A natural approach is to first project motion data into a structured semantic space and then train a generative model within that space. Such a paradigm ha…