Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision Paper • 2608.16812 • Published 20 days ago • 50
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture Paper • 2605.12500 • Published May 12 • 199
An Empirical Study of Training Pixel-Space Text-to-Image Diffusion Models Paper • 2608.16887 • Published 20 days ago • 35
LTX-2.5 Collection LTX-2.5 base models, quantized models and accompanying LoRAs and IC-LoRAs • 5 items • Updated 4 days ago • 56
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated 26 days ago • 107
Kimi K3 — De-Risked Collection 2.8T LatentMoE + KDA, refusal-surface reduced at the weight level (ABLITERATED). GGUF variants need llama.cpp PR #26185 — mainline cannot load them. • 3 items • Updated 10 days ago • 2
Inkling Collection Inkling is a versatile, customizable model that reasons over text, images, audio, with variable and efficient thinking effort. • 4 items • Updated Jul 27 • 62
MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset Paper • 2605.21272 • Published May 20 • 5