Tetris3D: 3D Scene Generation With Objects That Fit Together Paper • 2610.10539 • Published 1 day ago • 29
GRACE: Generation-aware latent compression for efficient video generation Paper • 2610.10524 • Published 1 day ago • 55
GRACE: Generation-aware latent compression for efficient video generation Paper • 2610.10524 • Published 1 day ago • 55
Foundations of Proactive Agents: Principles, Technical Layers, and Proactivity-Gym Paper • 2609.37267 • Published 9 days ago • 34
EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos Paper • 2609.39378 • Published 8 days ago • 70
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 7 days ago • 87
Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering Paper • 2609.38177 • Published 9 days ago • 72
Keep-or-Drop? Adaptive Tokenizer for Compact Video Representation Paper • 2608.24293 • Published Aug 25 • 14
DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models Paper • 2603.23499 • Published Mar 24 • 52
WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation Paper • 2603.16871 • Published Mar 17 • 61
Grounding World Simulation Models in a Real-World Metropolis Paper • 2603.15583 • Published Mar 16 • 156