← Search

Kolja Bauer

2 accepted papers

2026

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

CVPR 2026

Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scene dynamics, exploring multiple possible futures through full video synthesis remains prohibitively inefficient. We model scene dynamics orders of ma

Cited by 0SourcecodeScholar
2025

CleanDIFT: Diffusion Features without Noise

CVPR 2025poster

Internal features from large-scale pre-trained diffusion models have recently been established as powerful semantic descriptors for a wide range of downstream tasks. Works that use these features generally need to add noise to images before passing them through the model to obtain the semantic featu…