Divide and Conquer: Mixture-of-Bottleneck Experts in Informative Ordinal Space for Video-based Multimodal Sentiment Analysis
cs.MM, cs.CL
Submitted: 2026-09-16
Updated: 2026-09-16
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- End-to-end Semantic-centric Video-based Multimodal Affective Computing
- Gaussian Error Linear Units (GELUs)
- MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos
Related papers
- ControlFoley: Unified and Controllable Video-to-Audio Generation with Cross-Modal Conflict Handling
- Zero-shot Video Moment Retrieval via Off-the-shelf Multimodal Large Language Models
- Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
- A Rate-Distortion-Classification Approach for Lossy Image Compression