Rethinking Global Text Conditioning in Diffusion Transformers Paper • 2602.09268 • Published Feb 9 • 9
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation Paper • 2605.28091 • Published May 27 • 6
Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability Paper • 2606.28116 • Published Jun 26 • 1
view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • 28 days ago • 111
GAS: Improving Discretization of Diffusion ODEs via Generalized Adversarial Solver Paper • 2510.17699 • Published Oct 20, 2025 • 25
DiffusionBlocks: Blockwise Training for Generative Models via Score-Based Diffusion Paper • 2506.14202 • Published Jun 17, 2025 • 8
AlphaFlow: Understanding and Improving MeanFlow Models Paper • 2510.20771 • Published Oct 23, 2025 • 9
Elucidating the SNR-t Bias of Diffusion Probabilistic Models Paper • 2604.16044 • Published Apr 17 • 72
Seed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference Paper • 2508.02193 • Published Aug 4, 2025 • 139
Omni-Embed-Nemotron: A Unified Multimodal Retrieval Model for Text, Image, Audio, and Video Paper • 2510.03458 • Published Oct 3, 2025 • 4
LongCat-Audio-Codec: An Audio Tokenizer and Detokenizer Solution Designed for Speech Large Language Models Paper • 2510.15227 • Published Oct 17, 2025 • 4
DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation Paper • 2602.12160 • Published Feb 12 • 39