MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement Paper • 2610.11959 • Published 3 days ago • 70
TokenRouter: Efficient Serving System for Token-Level LLM Routing Paper • 2610.12242 • Published 3 days ago • 129
Beyond Selection: Token Parameterization for Extreme Visual Token Compression Paper • 2609.35232 • Published 13 days ago • 8
TabFM-Auto: Self-Evolving Pipelines for Tabular Foundation Models Paper • 2609.37989 • Published 12 days ago • 12
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR Paper • 2609.37868 • Published 12 days ago • 65
SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation Paper • 2609.36601 • Published 12 days ago • 95
Think Before You Score: Thinking Reward Model for Visual Generation Paper • 2609.37372 • Published 12 days ago • 104
MaLiang-Harness: A Programmable Path to Image and Video Generation Paper • 2609.34309 • Published 13 days ago • 394
bottlecapai/ThinkingCap-Qwen3.8-27B-NVFP4A4-AWQ Image-Text-to-Text • 20B • Updated 16 days ago • 1.46k • 12
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 24 days ago • 228
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published about 1 month ago • 266
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published Sep 10 • 657
Unlocking Lossless Speedups in LLMs via Discrete Diffusion Paper • 2609.04010 • Published Sep 3 • 114