TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 20 days ago • 14
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 24 days ago • 20
RoboProcessBench: Benchmarking Process-Aware Understanding in Vision-Language Robotic Manipulation Paper • 2606.13040 • Published Jun 11
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 31
TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 20 days ago • 14
TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 20 days ago • 14
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 24 days ago • 20
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 24 days ago • 20
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 31
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 31
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing Paper • 2602.02437 • Published Feb 2 • 80
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing Paper • 2602.12205 • Published Feb 13 • 83
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing Paper • 2602.02437 • Published Feb 2 • 80
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing Paper • 2602.12205 • Published Feb 13 • 83
ReMamber: Referring Image Segmentation with Mamba Twister Paper • 2403.17839 • Published Mar 26, 2024 • 1
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation Paper • 2309.00096 • Published Aug 31, 2023