papersTODAY 04:00 UTC
Open-UniMo Framework Unifies Motion-Language Understanding and Generation
A new arXiv paper introduces Open-UniMo, a framework that aims to combine human motion generation with motion understanding in a single model for open-world settings. The authors note that most existing motion-language models treat motion as a secondary modality attached to language, which limits how well they generalize. The work targets embodied AI systems that need to both produce and interpret human actions.