A Parameter-Free Few-Shot Evaluation for Elephant Vocalisation Classification
eess.AS, cs.LG, cs.SD, q-bio.QM
Submitted: 2026-08-14
Updated: 2026-09-23
Terminology
Sources
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
- Perch 2.0 transfers 'whale' to underwater tasks
- From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
- HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
- Perch 2.0: The Bittern Lesson for Bioacoustics
- Contrastive Learning of General-Purpose Audio Representations
Related papers
- X-VC: Zero-shot Streaming Voice Conversion in Codec Space
- Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers
- Anonymization, Not Elimination: Utility-Preserved Speech Anonymization
- Towards Audio Token Compression in Large Audio Language Models
- WaveScat: Wavelet Scattering Front-Ends with Self-Supervised Features for Speech Deepfake Detection
- ProPS: Prompted Profile Synthesis for Natural Language-Conditioned Speaker Embedding Distributions