ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 18 days ago • 215
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 15 days ago • 110
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 21 days ago • 85
PixelHacker: Image Inpainting with Structural and Semantic Consistency Paper • 2504.20438 • Published Apr 29, 2025 • 51
Running Featured 853 Agent Memory Leaderboard 🧠853 Unified memory evaluation · Results expected August 12.
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 28 days ago • 117
Atria Dawn: The Dawn of Agentic Superintelligence Paper • 2609.15818 • Published 18 days ago • 432
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 21 days ago • 265
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 18 days ago • 250
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 22 days ago • 173
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published 27 days ago • 166
Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models Paper • 2609.12641 • Published 21 days ago • 71
meta-llama/Llama-3.1-8B-Instruct Text Generation • 8B • Updated Sep 25, 2024 • 6.2M • • 8.14k
SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution Paper • 2609.05594 • Published 28 days ago • 36