arXiv:2508. 00307v4 Announce Type: replace-cross Abstract: We introduce a U-net model for 360{\deg} acoustic source localization formulated as a spherical semantic segmentation task.
By Belman Jahir Rodriguez, Sergio F. Chevtchenko, Marcelo Herrera Martinez, Yeshwanth Bethi, Saeed Afshar
arXiv:2605. 26310v2 Announce Type: replace Abstract: The detection of unmanned aerial vehicles (UAVs) is important for the protection of civilian and military infrastructure.
By Ungv\'ari Gerg\H{o}, Ferenc Braun, Attila \'Amon, P\'eter Kackst\"adter, J\'anos Volk, P\'eter Kov\'acs, Tam\'as D\'ozsa
arXiv:2607. 25887v1 Announce Type: cross Abstract: This paper explores the effectiveness of domain adaptation techniques when using convolutional neural network (CNN)-based and transformer-based feature representations for acoustic scene classification.
By Abhishek dileep, Shubham Sharma, Padmanabhan Rajan
arXiv:2606. 11922v1 Announce Type: cross Abstract: Recent respiratory sound classification (RSC) studies largely rely on CLS-token driven self-attention architectures such as the Audio Spectrogram Transformer (AST).
By Hemansh Shridhar, Miika Toikkanen, June-Woo Kim
arXiv:2608. 08911v1 Announce Type: cross Abstract: Recent advances in Passive Acoustic Monitoring (PAM) offer an opportunity to obtain ecological spatial point-process data at unprecedented scale.
By Jennifer N. Kampe, Changwoo J. Lee, Xin Shen, Ari Lehti\"o, Sandro von Brandenburg, Ossi Nokelainen, David B. Dunson, Otso Ovaskainen
arXiv:2511. 21325v2 Announce Type: replace-cross Abstract: Deepfake (DF) audio detectors still struggle to generalize to out of distribution inputs.
By Ido Nitzan Hidekel, Gal lifshitz, Khen Cohen, Dan Raviv