MiniMaxAI/MiniMax-H3 Image-Text-to-Video β’ 33B β’ Updated about 1 hour ago β’ 12.1k β’ β’ 2.68k
Scaling Properties of Text Conditioning in Visual Generation Paper β’ 2607.29679 β’ Published 7 days ago β’ 36
Mage Collection A family of lightweight multimodal models, including understanding and generation. β’ 8 items β’ Updated 11 days ago β’ 24
FastVideo/FastWan2.2-TI2V-5B-FullAttn-Diffusers Text-to-Video β’ 5B β’ Updated Nov 25, 2025 β’ 62.2k β’ 66
RLFR: Extending Reinforcement Learning for LLMs with Flow Environment Paper β’ 2510.10201 β’ Published Oct 11, 2025 β’ 37
Lumina Family Collection Lumina-T2X is a unified framework for Text to Any Modality Generation β’ 8 items β’ Updated Jul 30, 2024 β’ 6
stabilityai/stable-video-diffusion-img2vid-xt Image-to-Video β’ 2B β’ Updated Jul 10, 2024 β’ 173k β’ 3.37k