arXiv:2508. 00307v4 Announce Type: replace-cross Abstract: We introduce a U-net model for 360{\deg} acoustic source localization formulated as a spherical semantic segmentation task.
By Belman Jahir Rodriguez, Sergio F. Chevtchenko, Marcelo Herrera Martinez, Yeshwanth Bethi, Saeed Afshar
arXiv:2505.20961v2 Announce Type: replace-cross
Abstract: Sound source localization (SSL) is a critical technology for determining the position of sound sources in complex environments. However, exis...
By Yiyuan Yang, Shitong Xu, Niki Trigoni, Andrew Markham
arXiv:2609.35863v1 Announce Type: cross
Abstract: Modern bioacoustic foundation models like Perch and BirdNET can identify species with high discriminative accuracy, yet their confidence scores are o...
By Neha Sajja, Bart van Merri\"{e}nboer, Burcu Karagol Ayan, Tom Denton
arXiv:2608. 14287v1 Announce Type: cross Abstract: Passive acoustic sensing offers a critical, cost-efficient, and, crucially, passive alternative for detecting small unmanned aerial vehicles.
By Vadym Vilhurin, Volodymyr Sydorskyi, Andrii Shevtsov
arXiv:2609.15221v1 Announce Type: cross
Abstract: Passive acoustic monitoring can measure biodiversity at larger scales, but time--frequency annotation of animal vocalizations is expensive, site-spec...
By Tianyi Xu, Daniel Pimentel-Alarc\'on, Zuzana Bu\v{r}ivalov\'a, Claudia Sol\'is-Lemus
arXiv:2505. 18726v3 Announce Type: replace-cross Abstract: Can we determine someone's geographic location solely from the sounds they hear?
By Mustafa Chasmai, Wuao Liu, Subhransu Maji, Grant Van Horn
arXiv:2607. 14072v1 Announce Type: new Abstract: Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data.
By Mustafa Chasmai, Vincent Dumoulin, Jenny Hamer
The paper presents a lightweight ResNet-based two-stage cascade for passive acoustic monitoring of killer whales. First, it detects vocalizations, then it classifies confident detections into five eastern North Pacific ecotypes, abstaining on ambiguous calls. The pipeline achieves high macro‑F1 scores on the DCLDE 2027 dataset and improves real‑time inference speed, while active learning adapts the detector to new acoustic environments.
By Daniela Ruiz, Manuel Castellote, Zhongqi Miao, Carl Chalmers, Bruno Demuro, Rahul Dodhia, Pablo Arbelaez, Juan M. Lavista
The paper introduces MUSIC-Net, an end-to-end deep learning framework for near-field multi-user positioning that incorporates a two-stage MUSIC algorithm to isolate line-of-sight signal components and estimate surrogate distances. By embedding these MUSIC-derived objects into training, the method bypasses separate parameter estimation and path/source association, directly recovering user positions even in mixed LoS/NLoS multipath scenarios. Additionally, the authors employ split conformal prediction to provide statistically guaranteed confidence sets for each user’s position, achieving lower mean positioning error and tighter prediction regions compared to existing benchmarks.
By Jiaying Li, Haifeng Wen, Changsheng You, Yuanwei Liu, Hong Xing
Bioacoustic foundation models rely on large-scale citizen science platforms like Xeno-Canto for geographically and ecologically diverse data. Recent work has shown that supervision alone can produce SotA species detection models when trained on this large-scale data -- however, there remains unutilized potential in the form of recording metadata readily available within these community-driven data hubs.
The paper introduces a method for learning binaural sound localization by using egomotion as a supervisory signal. By tracking how a camera’s direction changes relative to a sound source during a video, the authors train an audio model to predict sound directions that align with visual estimates of camera motion derived from multi‑view geometry. They evaluate this approach on a newly proposed dataset of real‑world audio‑visual videos with egomotion, demonstrating that the model can learn from real data and perform well on sound localization tasks.
By Anna Min, Ziyang Chen, Hang Zhao, Andrew Owens
arXiv:2609.13281v1 Announce Type: cross
Abstract: Ocean-bottom seismometers (OBS), originally deployed for geophysical research, continuously record low-frequency sound for months to years across bro...
By Jocelyn Japnanto, Alex A. Saoulis, Miriam Romagosa, Rita Leit\~ao, Gabrielle Arrieta, M\'onica A. Silva, Matthew Graham, Ana M. G. Ferreira