MixiMotion: One-Step Text-to-Motion Generation via Asymmetric Set Distillation
cs.CV, cs.MM
Submitted: 2026-09-19
Updated: 2026-09-19
Terminology
Sources
- Qwen3-VL Technical Report
- MoFlow: One-Step Flow Matching for Human Trajectory Forecasting via Implicit Maximum Likelihood Estimation based Distillation
- Distilling the Knowledge in a Neural Network
- MotionPCM: Real-Time Motion Synthesis with Phased Consistency Model
- Implicit Maximum Likelihood Estimation
- SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation
- HY-Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation
- MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models