Publications
Structure From Tracking: Distilling Structure-Preserving Motion for Video Generation
Under Review
Distilling motion from a tracking model to preserve structure in generated videos.
Large Motion Video Autoencoding with Cross-modal Video VAE
ICCV 2025
Combining spatial and temporal compression with text guidance for high-fidelity video autoencoding.

