SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 12 days ago • 273
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 12 days ago • 694
Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision Paper • 2608.16812 • Published Aug 17 • 50
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture Paper • 2605.12500 • Published May 12 • 199
An Empirical Study of Training Pixel-Space Text-to-Image Diffusion Models Paper • 2608.16887 • Published Aug 17 • 36
LTX-2.5 Collection LTX-2.5 base models, quantized models and accompanying LoRAs and IC-LoRAs • 5 items • Updated 12 days ago • 65
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated Aug 10 • 109
Kimi K3 — De-Risked Collection 2.8T LatentMoE + KDA, refusal-surface reduced at the weight level (ABLITERATED). GGUF variants need llama.cpp PR #26185 — mainline cannot load them. • 2 items • Updated 10 days ago • 2
Inkling Collection Inkling is a versatile, customizable model that reasons over text, images, audio, with variable and efficient thinking effort. • 4 items • Updated Jul 27 • 65