LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Paper • 2608.30935 • Published 6 days ago • 29
Learning to Evaluate Before Improving: Automatic Rubric Induction for Automatic Research Agents Paper • 2608.31076 • Published 6 days ago • 20
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 6 days ago • 140
Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase Paper • 2608.29310 • Published 8 days ago • 27
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 6 days ago • 98
CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure Modes Paper • 2608.27455 • Published 10 days ago • 12
Thinking on Shots: Consistent Multi-Shot Video Editing with Agentic Reasoning Paper • 2608.26809 • Published 10 days ago • 7
What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents Paper • 2608.27260 • Published 10 days ago • 72
PAWBench: How Far Are We from Probabilistically Aligned World Modeling? Paper • 2608.27345 • Published 10 days ago • 144
Self-OPD: On-Policy Distillation for Flow Matching Models without Teacher Paper • 2608.26872 • Published 10 days ago • 82
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO Paper • 2608.27351 • Published 10 days ago • 22
TacForcing: Streaming Action Generation with Execution-Time Tactile Feedback Paper • 2608.25798 • Published 11 days ago • 8
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 11 days ago • 26
UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City Paper • 2608.27456 • Published 10 days ago • 112
Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL Paper • 2608.17253 • Published 18 days ago • 97
SecOPD: Mitigating Adaptive Prompt Injections by On-Policy Distillation Paper • 2608.21500 • Published 16 days ago • 41
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 11 days ago • 179
Better Retrieval, Worse Robustness:How Multi-hop RAG Amplifies Upstream ASR Errors Paper • 2608.22872 • Published 13 days ago • 7