PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published 13 days ago • 46
PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World Paper • 2605.13169 • Published May 13 • 21
GDRO: Group-level Reward Post-training Suitable for Diffusion Models Paper • 2601.02036 • Published Jan 5
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions Paper • 2506.23361 • Published Jun 29, 2025 • 1
MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives Paper • 2512.14699 • Published Dec 16, 2025 • 29
DiffDoctor: Diagnosing Image Diffusion Models Before Treating Paper • 2501.12382 • Published Jan 21, 2025
PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published 13 days ago • 46
PanoWorld: Towards Spatial Supersensing in 360^circ Panorama World Paper • 2605.13169 • Published May 13 • 21
PanoWorld: Towards Spatial Supersensing in 360^circ Panorama World Paper • 2605.13169 • Published May 13 • 21
Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items Paper • 2604.19748 • Published Apr 21 • 254
ShowUI-π: Flow-based Generative Models as GUI Dexterous Hands Paper • 2512.24965 • Published Dec 31, 2025 • 43
Robust-R1: Degradation-Aware Reasoning for Robust Visual Understanding Paper • 2512.17532 • Published Dec 19, 2025 • 68
Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and Editing Paper • 2512.17909 • Published Dec 19, 2025 • 37
Alchemist: Unlocking Efficiency in Text-to-Image Model Training via Meta-Gradient Data Selection Paper • 2512.16905 • Published Dec 18, 2025 • 32
TGDPO: Harnessing Token-Level Reward Guidance for Enhancing Direct Preference Optimization Paper • 2506.14574 • Published Jun 17, 2025 • 1
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds Paper • 2508.09454 • Published Aug 13, 2025