GLoRe: Detail-Preserving Text-Driven Two-Person Motion In-Betweening via Global Guidance and Local Refinement
Abstract
Motion in-betweening is fundamental to controllable character animation for virtual humans, games, and embodied interaction. While recent two-person motion in-betweening approaches have achieved good performance for generating coordinated interaction between two persons, they often result in degraded individual motion details. We introduce GLoRe, a novel three-stage detail-preserving framework containing an individual motion initialization stage that leverages a singleperson motion in-betweening model to generate detailed individual motions, a global guidance stage that provides global context for motion refinement, and a controlled local refinement stage that adaptively applies residual corrections. Specifically, in the global guidance stage, we propose a new global interaction condition that models sequence-level interaction demands from the initialized motions and input constraints; in the controlled local refinement stage, we introduce a new local residual gate that regulates the person- and frame-level corrections based on cross-person evidence. Extensive experiments on InterHuman and Inter-X demonstrate that our method achieves superior performance on text-driven twoperson motion in-betweening, effectively balancing individual motion quality and interaction coordination.
est. 32% chance this paper gets accepted at ICLR 2027.
What do you think this paper will get?
All positions stay anonymous.