Crystal Structure Prediction by Joint Equivariant Diffusion Paper • 2309.04475 • Published Jul 30, 2023
BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression Paper • 2606.10135 • Published Jun 8 • 2
LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows Paper • 2609.15863 • Published 1 day ago • 43
LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows Paper • 2609.15863 • Published 1 day ago • 43
The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping Paper • 2604.11297 • Published Apr 13 • 144
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing Paper • 2602.12205 • Published Feb 13 • 83
MOVA: Towards Scalable and Synchronized Video-Audio Generation Paper • 2602.08794 • Published Feb 9 • 159
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing Paper • 2602.02437 • Published Feb 2 • 80
ReVSeg: Incentivizing the Reasoning Chain for Video Segmentation with Reinforcement Learning Paper • 2512.02835 • Published Dec 2, 2025 • 10