SynCo: Learning Cross-Modal Synergy by Contrasting Interaction Residuals
cs.CV, cs.LG
Submitted: 2026-09-26
Updated: 2026-09-26
Terminology
Sources
- Gated Multimodal Units for Information Fusion
- A Cookbook of Self-Supervised Learning
- Orthogonalized Multimodal Contrastive Learning with Asymmetric Masking for Structured Representations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- What to align in multimodal contrastive learning?
- MultiBench: Multiscale Benchmarks for Multimodal Representation Learning
- Decoupled Weight Decay Regularization
- EarthScape: A Multimodal Dataset for Surficial Geologic Mapping and Earth Surface Analysis
- Representation Learning with Contrastive Predictive Coding
- InfMasking: Unleashing Synergistic Information by Contrastive Multimodal Interactions
- Nonnegative Decomposition of Multivariate Information
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models