Multi-Agent Egocentric World Model with Fine-Grained Embodied Interaction Paper • 2610.12299 • Published 4 days ago • 52
Tetris3D: 3D Scene Generation With Objects That Fit Together Paper • 2610.10539 • Published 5 days ago • 45
GRACE: Generation-aware latent compression for efficient video generation Paper • 2610.10524 • Published 5 days ago • 76
GRACE: Generation-aware latent compression for efficient video generation Paper • 2610.10524 • Published 5 days ago • 76
Foundations of Proactive Agents: Principles, Technical Layers, and Proactivity-Gym Paper • 2609.37267 • Published 13 days ago • 40
EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos Paper • 2609.39378 • Published 12 days ago • 74
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 11 days ago • 90
Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering Paper • 2609.38177 • Published 13 days ago • 74
Keep-or-Drop? Adaptive Tokenizer for Compact Video Representation Paper • 2608.24293 • Published Aug 25 • 14
DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models Paper • 2603.23499 • Published Mar 24 • 52
WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation Paper • 2603.16871 • Published Mar 17 • 61
Grounding World Simulation Models in a Real-World Metropolis Paper • 2603.15583 • Published Mar 16 • 156