DynamicsBoost: Dynamic Plausible Video Generation via Annotation-Free Continuation Preference Optimization
Despite significant progress in text-to-video generation, current models still suffer from unrealistic dynamics, temporal inconsistency, and unstable semantic alignment. Existing preference alignment approaches rely on costly and often ambiguous human or VLM-based video preference annotation, which