RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Paper • 2607.14187 • Published 7 days ago • 26
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 6 days ago • 183
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published 6 days ago • 159
X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Paper • 2607.12993 • Published 8 days ago • 131
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 8 days ago • 206
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation Paper • 2607.11886 • Published 9 days ago • 83
PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space Paper • 2607.05373 • Published 16 days ago • 65
SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History Paper • 2606.08671 • Published 29 days ago • 46
timaeus/rl-lm-pythia70m-formality-pos-beta0-grpo-nostd-gs4-tp1-tk0-pt80000-steerDotL12c2000s4-seed20 Updated 18 days ago • 1 • 1
MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision Paper • 2606.17162 • Published Jun 15 • 177