PhysVista: Benchmarking Physical Intelligence in VLMs via a Perception-Reasoning-Assessment Loop Paper • 2610.00559 • Published 12 days ago • 43
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 23 days ago • 43
Paragraph Boundaries Are Not White Space:Compression Depth as the Signature of Hierarchical Structure Paper • 2609.23551 • Published 22 days ago • 7
RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents Paper • 2609.25636 • Published 20 days ago • 13
Refinement Is Inherently Editable: Training-Free Prompt-to-Prompt Image Editing with Generative Refinement Network Paper • 2609.20633 • Published 25 days ago • 13
Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 25 days ago • 57
Calibrating Teacher--Student Discrepancy for On-Policy Distillation Paper • 2609.21619 • Published 24 days ago • 13
IntBMoE: Integrating Block-Level Conditioning into Expert Composition for Full-Participation Mixture-of-Experts Paper • 2609.21346 • Published 24 days ago • 101
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 24 days ago • 80
Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model Paper • 2609.18323 • Published 26 days ago • 120