SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Paper • 2607.21553 • Published 3 days ago • 21
view article Article Bringing Nunchaku 4-bit Diffusion Inference to Diffusers rootonchair, sayakpaul • 3 days ago • 43
view article Article Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers nvidia • 8 days ago • 79
PixelGen: Pixel Diffusion Beats Latent Diffusion with Perceptual Loss Paper • 2602.02493 • Published Feb 2 • 47
Alchemist: Turning Public Text-to-Image Data into Generative Gold Paper • 2505.19297 • Published May 25, 2025 • 86
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training Paper • 2501.04765 • Published Jan 8, 2025 • 1
REPA Works Until It Doesn't: Early-Stopped, Holistic Alignment Supercharges Diffusion Training Paper • 2505.16792 • Published May 22, 2025 • 1
What matters for Representation Alignment: Global Information or Spatial Structure? Paper • 2512.10794 • Published Dec 11, 2025 • 11
view article Article Profiling in PyTorch (Part 3): Attention is all you profile +2 ariG23498, sergiopaniego, sayakpaul, ror • 16 days ago • 38
Is Noise Conditioning Necessary for Denoising Generative Models? Paper • 2502.13129 • Published Feb 18, 2025 • 2
Back to Basics: Let Denoising Generative Models Denoise Paper • 2511.13720 • Published Nov 17, 2025 • 71
Scaling Rectified Flow Transformers for High-Resolution Image Synthesis Paper • 2403.03206 • Published Mar 5, 2024 • 72
PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis Paper • 2310.00426 • Published Sep 30, 2023 • 63
DC-AE 1.5: Accelerating Diffusion Model Convergence with Structured Latent Space Paper • 2508.00413 • Published Aug 1, 2025 • 7
LTX-2.3 Creative Lab Collection LoRAs and IC-LoRAs, trained on the LTX-2.3 model • 25 items • Updated 7 days ago • 83