OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.32M • 1.08k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 2.85M • • 2k uv-scripts/ocr Updated 24 days ago • 2.98k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 216k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 3.4k • 737 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 107 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 40.4k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 9 days ago • 2.09M • • 12.2k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 120k • 2.46k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 5 days ago • 19.2k • 814 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated 13 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 94 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 35.8k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 963
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 35.8k • • 2.94k
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.28k • 67 Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 17 StuPASE Speech Enhancement 🎙 17 Studio-quality generative speech enhancement
Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 123k • • 2.49k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated about 23 hours ago • 954 • 97
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 19 days ago • 48
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 19 days ago • 48
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.28k • 67 Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 17 StuPASE Speech Enhancement 🎙 17 Studio-quality generative speech enhancement
Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.32M • 1.08k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 2.85M • • 2k uv-scripts/ocr Updated 24 days ago • 2.98k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 216k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 3.4k • 737 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 107 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 40.4k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 9 days ago • 2.09M • • 12.2k
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 123k • • 2.49k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated about 23 hours ago • 954 • 97
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 120k • 2.46k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 5 days ago • 19.2k • 814 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 19 days ago • 48
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 19 days ago • 48
Model training merve/smol-vision Image-Text-to-Text • Updated 13 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 94 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 35.8k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 963
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 35.8k • • 2.94k