view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • 1 day ago • 69
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs Paper • 2608.03573 • Published 6 days ago • 45
Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning Paper • 2608.03571 • Published 6 days ago • 40
Qwen3.8 27B Compressions and GGUFs Collection Collection of Qwen3.8 27B Multimodal Compressions • 2 items • Updated 1 day ago • 1
VideoGuard Multimodal Collection Collection of VideoGuard Reasoning Models • 4 items • Updated about 1 hour ago • 1
view article Article Training a coding agent using the OpenCode harness in remote HF sandboxes with TRL and OpenEnv sergiopaniego • 6 days ago • 19
view article Article LFM2.5-Encoders for Fast Long-Context Inference on CPU LiquidAI • 14 days ago • 68
view article Article The OlmoEarth Platform: Geospatial inference at planetary scale allenai • 14 days ago • 40
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper • 2607.28609 • Published 13 days ago • 68
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published 6 days ago • 90
Opus Qwen3.5 Distilled Edge — Device Models Collection Collection of Opus Distilled (Qwen3.5) • 4 items • Updated 1 day ago • 1
view article Article Hugging Face on AMD Instinct MI455X: First Transformers Results badaoui • 19 days ago • 18
view article Article mDenseOn with the mLateOn: Open Multilingual, Long-Context, and Code Retrieval Models lightonai • 12 days ago • 35
oMEGA Collection Spatial Reasoning with Concise Notes for Vision Tasks • 2 items • Updated 4 days ago • 1
OpusGLM SFT Collection Collection of Qwen3.5 Multimodal Models Fine-Tuned with OpusGLM Traces • 4 items • Updated 4 days ago • 1
Explicit Layer Modeling for Video Object Insertion and Layer Decomposition Paper • 2607.25802 • Published 15 days ago • 8
CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition Paper • 2607.25294 • Published 15 days ago • 49
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 14 days ago • 139
Common Audio Classification (latest) Collection Collection of Audio Segmentation Models • 2 items • Updated 4 days ago • 1