SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science? Paper • 2608.19799 • Published Aug 20 • 67
SRPO Collection Official Collections for SRPO: Self-Referential Policy Optimization for Vision-Language-Action Models, including SFT and RL models. • 6 items • Updated Aug 6 • 2
RoboOmni Collection Proactive Robot Manipulation in Omni-modal Context • 9 items • Updated Jul 11 • 14
MHA2MLA Collection The MHA2MLA model published in the paper "Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-Based LLMs" • 22 items • Updated Jul 11 • 3
Game-RL Collection [ICLR 2026] Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning • 8 items • Updated Jul 11 • 6
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Paper • 2607.29613 • Published Jul 31 • 29
World-aware Planning Narratives Enhance Large Vision-Language Model Planner Paper • 2506.21230 • Published Jun 26, 2025 • 1
LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models Paper • 2510.13626 • Published Oct 15, 2025 • 48
RoboOmni: Proactive Robot Manipulation in Omni-modal Context Paper • 2510.23763 • Published Oct 27, 2025 • 62