Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification
eess.AS, cs.AI, cs.SD
Submitted: 2026-07-15
Updated: 2026-09-18
Comments: Accepted to DCASE Workshop 2026, github repo "https://github.com/TioSisai/mismatch-weighted-facility-location"
Code: https://github.com/TioSisai/mismatch-weighted-facility-location
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Large sample analysis of the median heuristic
- UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction
Related papers
- X-VC: Zero-shot Streaming Voice Conversion in Codec Space
- Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers
- Anonymization, Not Elimination: Utility-Preserved Speech Anonymization
- Towards Audio Token Compression in Large Audio Language Models
- WaveScat: Wavelet Scattering Front-Ends with Self-Supervised Features for Speech Deepfake Detection
- ProPS: Prompted Profile Synthesis for Natural Language-Conditioned Speaker Embedding Distributions