Self-Supervised Learning of Structured Dynamics from Videos Paper • 2607.21576 • Published 16 days ago • 19
Self-Supervised Learning of Structured Dynamics from Videos Paper • 2607.21576 • Published 16 days ago • 19
Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers Paper • 2606.32020 • Published Jun 30 • 4
Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers Paper • 2606.32020 • Published Jun 30 • 4
PEEK: Picking Essential frames via Efficient Knowledge distillation Paper • 2605.31029 • Published May 29 • 20
An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning Paper • 2211.16780 • Published Apr 16 • 3
An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning Paper • 2211.16780 • Published Apr 16 • 3
LeGrad: An Explainability Method for Vision Transformers via Feature Formation Sensitivity Paper • 2404.03214 • Published Apr 4, 2024 • 3
VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs Paper • 2512.21194 • Published Dec 24, 2025
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model Paper • 2603.26357 • Published Mar 27 • 4
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model Paper • 2603.26357 • Published Mar 27 • 4
Flow Straighter and Faster: Efficient One-Step Generative Modeling via MeanFlow on Rectified Trajectories Paper • 2511.23342 • Published Nov 28, 2025 • 15
CoPE-VideoLM: Codec Primitives For Efficient Video Language Models Paper • 2602.13191 • Published Feb 13 • 32
AMoE: Agglomerative Mixture-of-Experts Vision Foundation Model Paper • 2512.20157 • Published Dec 23, 2025 • 5
YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation Paper • 2601.08441 • Published Jan 13 • 8
YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation Paper • 2601.08441 • Published Jan 13 • 8