SGMD: Score Gradient Matching Distillation for Few-Step Video Diffusion Distillation Paper • 2605.30116 • Published May 28 • 4
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Paper • 2607.21553 • Published 5 days ago • 35
Xiaomi-Robotics-U0 Collection Unified embodied synthesis model that bridges foundation image generation and embodied world modeling • 2 items • Updated 15 days ago • 10
Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget Paper • 2607.13125 • Published 10 days ago • 137
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence Paper • 2607.07675 • Published 20 days ago • 64
KVAE 2.0 Collection KVAE 2.0 is a family of video tokenizers with a time compression ratio of 4 and spacial compression ratio of 8 and 16 • 2 items • Updated Apr 16 • 5
KVAE-Audio Collection KVAE-Audio is a continuous full-band audio waveform autoencoder • 1 item • Updated 29 days ago • 7
LTX-2.3 Creative Lab Collection LoRAs and IC-LoRAs, trained on the LTX-2.3 model • 25 items • Updated 9 days ago • 84
AAD-1: Asymmetric Adversarial Distillation for One-Step Autoregressive Video Generation Paper • 2606.03972 • Published Jun 2 • 14
From Pixels to Words -- Towards Native One-Vision Models at Scale Paper • 2605.28820 • Published May 27 • 76
RTDMD Collection Reinforcing Few-step Generators via Reward-Tilted Distribution Matching • 5 items • Updated Jun 2 • 3
Reinforcing Few-step Generators via Reward-Tilted Distribution Matching Paper • 2605.26108 • Published May 25 • 7
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion Paper • 2605.23902 • Published May 22 • 47