TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 3 days ago • 10
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 7 days ago • 17
RoboProcessBench: Benchmarking Process-Aware Understanding in Vision-Language Robotic Manipulation Paper • 2606.13040 • Published Jun 11
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 30
TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 3 days ago • 10
TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation Paper • 2608.16765 • Published 3 days ago • 10
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 7 days ago • 17
H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models Paper • 2608.13049 • Published 7 days ago • 17
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 30
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Paper • 2606.11683 • Published Jun 10 • 30
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing Paper • 2602.02437 • Published Feb 2 • 80
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing Paper • 2602.12205 • Published Feb 13 • 83
UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing Paper • 2602.02437 • Published Feb 2 • 80
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing Paper • 2602.12205 • Published Feb 13 • 83
ReMamber: Referring Image Segmentation with Mamba Twister Paper • 2403.17839 • Published Mar 26, 2024 • 1
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation Paper • 2309.00096 • Published Aug 31, 2023