Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published 4 days ago • 34
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Paper • 2606.28128 • Published Jun 26 • 53
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Paper • 2606.28128 • Published Jun 26 • 53
HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining Paper • 2606.20521 • Published Jun 18 • 14
HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining Paper • 2606.20521 • Published Jun 18 • 14
Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars Paper • 2602.01538 • Published Feb 2 • 15
ConsistEdit: Highly Consistent and Precise Training-free Visual Editing Paper • 2510.17803 • Published Oct 20, 2025 • 14
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence Paper • 2509.12203 • Published Sep 15, 2025 • 20
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence Paper • 2509.12203 • Published Sep 15, 2025 • 20
Progressive Disentangled Representation Learning for Fine-Grained Controllable Talking Head Synthesis Paper • 2211.14506 • Published Nov 26, 2022 • 1
Progressive Disentangled Representation Learning for Fine-Grained Controllable Talking Head Synthesis Paper • 2211.14506 • Published Nov 26, 2022 • 1