Qwen-Image-Flash: Rethinking the Training Recipe for Few-Step Distillation
cs.CV, cs.AI, cs.GR, cs.LG
Submitted: 2026-06-02
Updated: 2026-09-14
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Optimizing Few-Step Generation with Adaptive Matching Distillation
- Flow-OPD: On-Policy Distillation for Flow Matching Models
- Mean Flows for One-step Generative Modeling
- Distribution Matching Distillation Meets Reinforcement Learning
- DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models
- Decoupled DMD: CFG Augmentation as the Spear, Distribution Matching as the Shield
- ERNIE-Image Technical Report
- Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
- TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
- Wan-Image: Pushing the Boundaries of Generative Visual Intelligence
- Progressive Distillation for Fast Sampling of Diffusion Models
- JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation
- TIIF-Bench: How Does Your T2I Model Follow Your Instructions?
- Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis
- MiMo-V2-Flash Technical Report
- Qwen3 Technical Report
- Qwen-Image-2.0 Technical Report
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models