PRIMARY SOURCE
MoSAT: Human Motion Generation from Spatial Audio and Textual Description
2 days, 10 hours ago
Read source
arXiv cs.GR details MoSAT, a motion-synthesis model that uses spatial audio plus text to drive full-body animation. The team pairs it with STAM, a new dataset of motion sequences, directional sound, and rich annotations, aiming for more precise and temporally coherent character responses.
With your permission, GameDev.net uses analytics cookies to understand how people use the platform. You can accept analytics or continue with necessary cookies only. Learn more