Efficient Text-to-Image Generation: An Adaptive Step Schedule Controller for Diffusion Models
cs.CV, cs.AI
Submitted: 2026-09-15
Updated: 2026-09-15
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
- Hierarchical Text-Conditional Image Generation with CLIP Latents
- Knowledge Distillation in Iterative Generative Models for Improved Sampling Speed
- BK-SDM: A Lightweight, Fast, and Cheap Version of Stable Diffusion
- DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models
- CLIPScore: A Reference-free Evaluation Metric for Image Captioning
- Denoising Diffusion Implicit Models
- Pseudo Numerical Methods for Diffusion Models on Manifolds
- Denoising Diffusion Step-aware Models
- Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis
- Classifier-Free Diffusion Guidance
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models