Marionette: Predicting World States, Rendering Geometry, Painting Appearance Paper • 2608.14530 • Published 16 days ago • 33
HelloWorld: Enabling Socially Interactive Characters in Video World Models Paper • 2608.05070 • Published 25 days ago • 40
LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models Paper • 2607.08770 • Published Jul 9 • 37
BRDFusion: Physics Meets Generation for Urban Scene Inverse Rendering Paper • 2606.17049 • Published Jun 15 • 28
Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models Paper • 2606.12412 • Published Jun 10 • 21