2026
Unifying Precise Keyframes and Semantic Control via Multi-level Diffusion
CVPR 2026
Text-conditioned human motion in-betweening leverages keyframes for spatio-temporal control, with text providing high-level semantic guidance for the transitions. However, existing methods are unable to establish a coherent alignment between textual semantics and the spatio-temporal constraints prov