Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories Paper • 2607.15330 • Published 19 days ago • 71
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 19 days ago • 206
Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process Paper • 2607.03748 • Published about 1 month ago • 41
TACO: Tool-Augmented Credit Optimization for Agentic Tool Use Paper • 2606.30251 • Published Jun 29 • 22
ActiveMimic: Egocentric Video Pretraining with Active Perception Paper • 2606.06194 • Published Jun 4 • 2
Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization Paper • 2605.28109 • Published May 27 • 23
ankile/real01c-insert-marker-d1-baseline-uniform-r1-25k-nocf-eval-sobol20 Viewer • Updated May 28 • 5.22k • 34 • 1