MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Paper • 2607.14189 • Published 27 days ago • 34
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation Paper • 2607.14202 • Published 27 days ago • 42
AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation Paper • 2605.13724 • Published May 13 • 105
Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling Paper • 2604.05072 • Published Apr 10 • 18
VidEmo: Affective-Tree Reasoning for Emotion-Centric Video Foundation Models Paper • 2511.02712 • Published Nov 4, 2025 • 6
MODA: MOdular Duplex Attention for Multimodal Perception, Cognition, and Emotion Understanding Paper • 2507.04635 • Published Jul 7, 2025 • 1