Motion-Omni Makes Dialogue Models Decide What to Say and How to Move
Motion-Omni jointly generates speech, facial expression, hand gestures, and full-body motion from shared hidden states. The system approaches a teacher cascade in motion quality while delivering faster-than-real-time responses.
Read more