WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation Paper • 2608.24479 • Published 2 days ago • 45
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 15 days ago • 288
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 17 days ago • 341
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 28 days ago • 309
UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models Paper • 2607.23373 • Published Jul 25 • 7
Search and Refine During Think: Autonomous Retrieval-Augmented Reasoning of LLMs Paper • 2505.11277 • Published May 16, 2025 • 64
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 27.2k • • 2.94k
anorim/bertimbau-fusion-2-dareties-p0.9-k0.5-l1-bestcross-hatebr-hspt-olidbr-toldbr-tupy-v7 0.1B • Updated Jul 17 • 8 • 1
ABot-N1: Toward a General Visual Language Navigation Foundation Model Paper • 2607.10383 • Published Jul 14 • 102
When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search Paper • 2606.27669 • Published Jun 26 • 16