P. Biswal and P. K. Mohanty, “Development of quadruped walking robots: A review,” Ain Shams Eng. J. 12, 2017-2031 (2021).
10.1016/j.asej.2020.11.005Z. Cao, B. Nie, Y. Zhang, and Y. Gao, “Minimizing acoustic noise: Enhancing quiet locomotion for quadruped robots in indoor applications,” Proc. IROS, 17972-17979 (2025).
10.1109/IROS60139.2025.11246563G. Ince, K. Nakadai, T. Rodemann, H. Tsujino, and J.-I. Imura, “Ego noise cancellation of a robot using missing feature masks,” Appl. Intell. 34, 360-371 (2011).
10.1007/s10489-011-0285-0K. Furukawa, K. Okutani, K. Nagira, T. Otsuka, K. Itoyama, K. Nakadai, and H. G. Okuno, “Noise correlation matrix estimation for improving sound source localization by multirotor UAV,” Proc. IROS, 3943-3948 (2013).
10.1109/IROS.2013.6696920A. Briegleb, A. Schmidt, and W. Kellermann, “Deep clustering for single-channel ego-noise suppression,” Proc. ICA, 2813-2820 (2019).
D. Mukhutdinov, A. Alex, A. Cavallaro, and L. Wang, “Deep learning models for single-channel speech enhancement on drones,” IEEE Access, 11, 22993-23007 (2023).
10.1109/ACCESS.2023.3253719H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet: A low complexity speech enhancement framework for full-band audio based on deep filtering,” Proc. ICASSP, 7407-7411 (2022).
10.1109/ICASSP43922.2022.9747055H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet2: Towards real-time speech enhancement on embedded devices for full-band audio,” Proc. IWAENC, 1-5 (2022).
10.1109/IWAENC53105.2022.9914782H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet: Perceptually motivated real-time speech enhancement,” Proc. Interspeech, 2008-2009 (2023).
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “An algorithm for intelligibility prediction of time- frequency weighted noisy speech,” IEEE Trans. Audio Speech Lang. Process. 19, 2125-2136 (2011).
10.1109/TASL.2011.2114881ITU-T Rec. P.862, Perceptual Evaluation of Speech Quality (PESQ): An Objective Method for End-to-end Speech Quality Assessment of Narrow-band Telephone Networks and Speech Codecs, 2001.
J. Le Roux, S. Wisdom, H. Erdogan, and J. R. Hershey, “SDR – half-baked or well done?,” Proc. ICASSP, 626-630 (2019).
10.1109/ICASSP.2019.8683855Acoustic Camera - SoundCam 2.0, https://www.cae-systems.de/en/products/acoustic-camera-sound-source-localization/soundcam-20.html, (Last viewed August 5, 2026).
Recording Mode,https://izyrec.zendesk.com/hc/en-us/articles/7443876632335-Recording-Mode, (Last viewed August 5, 2026).
Galaxy S24 Ultra Specifications, https://www.samsung.com/sec/smartphones/galaxy-s24-ultra/specs/, (Last viewed August 5, 2026).
Conversational Speech Dataset for Emotion Classification, https://aihub.or.kr/aihubdata/data/view.do?dataSetSn=263, (Last viewed July 9, 2026).
J.-U. Bang, S. Yun, S.-H. Kim, M.-Y. Choi, M.-K. Lee, Y.-J. Kim, D.-H. Kim, J. Park, Y.-J. Lee, and S.-H. Kim, “KsponSpeech: Korean spontaneous speech corpus for automatic speech recognition,” Appl. Sci. 10, 6936 (2020).
10.3390/app10196936Noisy Speech Database for Training Speech Enhancement Algorithms and TTS Models, https://doi.org/10.7488/ds/2117, (Last viewed August 5, 2026).
10.7488/ds/2117C. K. A. Reddy, V. Gopal, R. Cutler, E. Beyrami, R. Cheng, H. Dubey, S. Matusevych, R. Aichner, A. Aazami, S. Braun, P. Rana, S. Srinivasan, and J. Gehrke, “The INTERSPEECH 2020 deep noise suppression challenge: Datasets, subjective testing framework, and challenge results,” Proc. Interspeech, 2492-2496 (2020).
10.21437/Interspeech.2020-3038- Publisher :The Acoustical Society of Korea
- Publisher(Ko) :한국음향학회
- Journal Title :The Journal of the Acoustical Society of Korea
- Journal Title(Ko) :한국음향학회지
- Volume : 45
- No :5
- Pages :572-581
- Received Date : 2026-07-14
- Revised Date : 2026-08-11
- Accepted Date : 2026-08-24
- DOI :https://doi.org/10.7776/ASK.2026.45.5.572



The Journal of the Acoustical Society of Korea









