Sound Event Detection

Momentum

9 papers in the last four weeks, against 1 the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 41

All topics
CardsList
  1. CARES: A Controlled Synthetic Benchmark of Speaker Reactions to Sound

    Oct 7, 2026Marcel Gibier, Thomas Thebaud, Olivier Boëffard +1Audio UnderstandingAudio Captioning

  2. Logbook: Extremely Long-form Audio Event Understanding

    Oct 5, 2026Kwanghee Choi, Suwon Shon, Dmitriy Serdyuk +7Audio-Language Model EvaluationAudio Understanding

  3. Bad: Taming the Bioacoustic Data Deluge with a Bat Activity Detector

    Sep 29, 2026Stefano Ciapponi, Santiago Martinez Balvanera, Andrea Cesaretti +2Sound Event DetectionBioacoustics

  4. SincDPNet: Interpretable Raw-Waveform Bathroom Activity Recognition for Assistive Living

    Sep 28, 2026Debolina Chowdhury, Suman Samui, Sujoy SahaTiny MLSound Event Detection

  5. OpenWhistle: A Large-Scale Longitudinal Dataset and Benchmark of Bottlenose Dolphin Vocalizations

    Sep 28, 2026Faadil Mustun, Chiara Semenzin, Roberto Dessi +7Sound Event DetectionBioacoustics

  6. SAIL: Spatial Audio Intelligence with Large Language Models via Disentangled Acoustic-Spatial Encoding and Dual-Stream Q-Former

    Sep 28, 2026Zhengding Luo, Jinyang Wu, Haozhe Ma +3Audio ReasoningDirection-of-Arrival Estimation

  7. REVE: Efficient Hallucination Correction for Large Audio-Language Models via Reused Encoder States

    Sep 22, 2026Hongjin Song, Jiasheng Kuang, Xinyu Yang +4Audio-Language Model Hallucination DetectionSound Event Detection

  8. Misrecognition or Abstraction? Rethinking Outputs of Sound Event Recognition

    Sep 20, 2026Naoya Tomida, Yuki Okamoto, Keisuke ImotoSound Event DetectionAudio Classification

  9. Augmenting Large Audio-Language Models with Frame-Level Grounding for Fine-Grained Temporal Perception

    Sep 14, 2026Yanfeng Shi, Yan Song, Junhui Li +4Audio UnderstandingAudio-Language Models

  10. NVV-Locator: From Transcript Tags to Acoustic Boundaries for Fine-Grained Nonverbal Vocalization Grounding

    Sep 9, 2026Yuang Cao, Bingshen Mu, Zhennan Lin +7Audio UnderstandingSound Event Detection

  11. Efficient Passive Acoustic Monitoring of Killer Whales Using a Two-Stage Detection and Ecotype Classification Cascade

    Sep 1, 2026Daniela Ruiz, Manuel Castellote, Zhongqi Miao +5Sound Event DetectionBioacoustics

  12. Smartphone Audio Based Distress Detection

    Aug 4, 2026Anil Sharma, Sarthak Ahuja, Mayank Gautam +1Sound Event DetectionAudio Classification

  13. An End-to-End Workflow for Fin Whale Song Detection, Note Characterization, and Localization with Distributed Acoustic Sensing

    Aug 3, 2026Dídac Diego-Tortosa, Miriam Romagosa, Arantza Ugalde +4Distributed Acoustic SensingSound Event Detection

  14. Ultra-Compact CNN Architectures for Tropical Bird Audio Detection on Microcontrollers

    Jul 22, 2026Muhammad Mun'im Ahmad Zabidi, Mohd Yamani Idna Idris, Norisma IdrisTiny MLSound Event Detection

  15. RealDESED: A Real-World Domestic Sound Event Detection Benchmark

    Jul 18, 2026Florian Schmid, Paul Primus, Alexander Fichtinger +3Sound Event Detection

  16. Can Tokens Compete? Token Representations against Supervised CNN Backbones for BirdCLEF+ 2026

    Jul 16, 2026Anthony Miyaguchi, Murilo Gustineli, Adrian CheungAudio Representation LearningNeural Audio Codecs

  17. Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification

    Jul 15, 2026Shiqi Zhang, Tuomas VirtanenMulti-Label ClassificationActive Learning

  18. Semi-Supervised Sound Event Detection with Conditional Mixup and Embedding-Level Contrastive Loss

    Jun 29, 2026Nian Shao, Xian Li, Xiaofei LiContrastive LearningSound Event Detection

  19. From General-Purpose Audio Tagging to Spatially Grounded Sound Event Localization and Detection

    Jun 26, 2026Stefano Giacomelli, Stefano Damiano, Claudia Rinaldi +2Audio Representation LearningDirection-of-Arrival Estimation

  20. Soroll-IA: A Weakly Labeled Audio Dataset for Real-World Industrial Port Monitoring

    Jun 24, 2026Javier Naranjo-Alcazar, Jordi Grau-Haro, Ruben Ribes-Serrano +2Sound Event DetectionWeakly Supervised Learning

  21. An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

    Jun 20, 2026Corrado Baccheschi, Patrizio DazziEfficient Neural Network InferenceReservoir Computing

  22. A Neuromorphic Trigger for Efficient Audio Event Detection

    Jun 16, 2026Benjamin Hatton, Oliver Rhodes, Luca PeresEfficient Neural Network InferenceSpiking Neural Networks

  23. Dolph2Vec: Self-Supervised Representations of Dolphin Vocalizations

    Jun 10, 2026Chiara Semenzin, Faadil Mustun, Roberto Dessi +5Audio Representation LearningAudio Understanding

  24. Time-frequency localization of bird calls in dense soundscapes

    Jun 9, 2026Simen Hexeberg, Fanghui Tong, Hari Vishnu +1Sound Event DetectionObject Detection

  25. SagnacAssisted Enhanced OTDR for Distributed Acoustic Sensing: A Standardized Benchmark and Engineering Evaluation Framework

    Jun 4, 2026Weiguang Wang, Fugen Wu, Hailing Wang +4Distributed Acoustic SensingSound Event Detection

  26. Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier

    May 22, 2026Martynas Dumpis, Tuomas VirtanenGradient-Based AttributionIntegrated Gradients

  27. CoarseSoundNet: Building a reliable model for ecological soundscape analysis

    May 20, 2026Alexander Gebhard, Andreas Triantafyllopoulos, Dominik Arend +4Sound Event DetectionPassive Acoustic Monitoring