Incentive Noise and Structural Prior Infusion for Multi-modal Object Re-Identification
cs.CV
Submitted: 2026-09-21
Updated: 2026-09-21
Code: https://github.com/zw-absin/INSPI
Terminology
Sources
- UniCat: Crafting a Stronger Fusion Baseline for Multimodal Re-Identification
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- In Defense of the Triplet Loss for Person Re-Identification
- Mixture of Noise for Pre-Trained Model-Based Class-Incremental Learning
- NEXT: Multi-Grained Mixture of Experts via Text-Modulation for Multi-Modal Object Re-Identification
- ICPL-ReID: Identity-Conditional Prompt Learning for Multi-Spectral Object Re-Identification
- Causal Bootstrapped Alignment for Unsupervised Video-Based Visible-Infrared Person Re-Identification
- Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
- Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free
- Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer
- DINOv3
- Reliable Multi-Modal Object Re-Identification via Modality-Aware Graph Reasoning
- GraFT: Gradual Fusion Transformer for Multimodal Re-Identification
- Dynamic Enhancement Network for Partial Multi-modality Person Re-identification
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models