ZoomDiff: A High-Fidelity Diffusion Model for Dual-Camera Smooth Zooming
cs.CV
Submitted: 2026-09-23
Updated: 2026-09-23
Project page: https://jiayi-hit.github.io/ZoomDiff.github.io
Terminology
Sources
- Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
- VideoCrafter1: Open Diffusion Models for High-Quality Video Generation
- Adam: A Method for Stochastic Optimization
- CAT: Cross Attention in Vision Transformer
- Learning Transferable Visual Models From Natural Language Supervision
- Towards Accurate Generative Models of Video: A New Metric & Challenges
- Wan: Open and Advanced Large-Scale Video Generative Models
- Framer: Interactive Frame Interpolation
- Generative Inbetweening: Adapting Image-to-Video Models for Keyframe Interpolation
- ViBiDSampler: Enhancing Video Interpolation Using Bidirectional Diffusion Sampler
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models