All Issue

2026 Vol.45, Issue 5 Preview Page
30 September 2026. pp. 572-581
Abstract
References
1

P. Biswal and P. K. Mohanty, “Development of quadruped walking robots: A review,” Ain Shams Eng. J. 12, 2017-2031 (2021).

10.1016/j.asej.2020.11.005
2

Z. Cao, B. Nie, Y. Zhang, and Y. Gao, “Minimizing acoustic noise: Enhancing quiet locomotion for quadruped robots in indoor applications,” Proc. IROS, 17972-17979 (2025).

10.1109/IROS60139.2025.11246563
3

G. Ince, K. Nakadai, T. Rodemann, H. Tsujino, and J.-I. Imura, “Ego noise cancellation of a robot using missing feature masks,” Appl. Intell. 34, 360-371 (2011).

10.1007/s10489-011-0285-0
4

K. Furukawa, K. Okutani, K. Nagira, T. Otsuka, K. Itoyama, K. Nakadai, and H. G. Okuno, “Noise correlation matrix estimation for improving sound source localization by multirotor UAV,” Proc. IROS, 3943-3948 (2013).

10.1109/IROS.2013.6696920
5

A. Briegleb, A. Schmidt, and W. Kellermann, “Deep clustering for single-channel ego-noise suppression,” Proc. ICA, 2813-2820 (2019).

6

D. Mukhutdinov, A. Alex, A. Cavallaro, and L. Wang, “Deep learning models for single-channel speech enhancement on drones,” IEEE Access, 11, 22993-23007 (2023).

10.1109/ACCESS.2023.3253719
7

H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet: A low complexity speech enhancement framework for full-band audio based on deep filtering,” Proc. ICASSP, 7407-7411 (2022).

10.1109/ICASSP43922.2022.9747055
8

H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet2: Towards real-time speech enhancement on embedded devices for full-band audio,” Proc. IWAENC, 1-5 (2022).

10.1109/IWAENC53105.2022.9914782
9

H. Schröter, A. N. Escalante-B., T. Rosenkranz, and A. Maier, “DeepFilterNet: Perceptually motivated real-time speech enhancement,” Proc. Interspeech, 2008-2009 (2023).

10

C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “An algorithm for intelligibility prediction of time- frequency weighted noisy speech,” IEEE Trans. Audio Speech Lang. Process. 19, 2125-2136 (2011).

10.1109/TASL.2011.2114881
11

ITU-T Rec. P.862, Perceptual Evaluation of Speech Quality (PESQ): An Objective Method for End-to-end Speech Quality Assessment of Narrow-band Telephone Networks and Speech Codecs, 2001.

12

J. Le Roux, S. Wisdom, H. Erdogan, and J. R. Hershey, “SDR – half-baked or well done?,” Proc. ICASSP, 626-630 (2019).

10.1109/ICASSP.2019.8683855
13
15

Galaxy S24 Ultra Specifications, https://www.samsung.com/sec/smartphones/galaxy-s24-ultra/specs/, (Last viewed August 5, 2026).

16

Conversational Speech Dataset for Emotion Classification, https://aihub.or.kr/aihubdata/data/view.do?dataSetSn=263, (Last viewed July 9, 2026).

17

J.-U. Bang, S. Yun, S.-H. Kim, M.-Y. Choi, M.-K. Lee, Y.-J. Kim, D.-H. Kim, J. Park, Y.-J. Lee, and S.-H. Kim, “KsponSpeech: Korean spontaneous speech corpus for automatic speech recognition,” Appl. Sci. 10, 6936 (2020).

10.3390/app10196936
18

Noisy Speech Database for Training Speech Enhancement Algorithms and TTS Models, https://doi.org/10.7488/ds/2117, (Last viewed August 5, 2026).

10.7488/ds/2117
19

C. K. A. Reddy, V. Gopal, R. Cutler, E. Beyrami, R. Cheng, H. Dubey, S. Matusevych, R. Aichner, A. Aazami, S. Braun, P. Rana, S. Srinivasan, and J. Gehrke, “The INTERSPEECH 2020 deep noise suppression challenge: Datasets, subjective testing framework, and challenge results,” Proc. Interspeech, 2492-2496 (2020).

10.21437/Interspeech.2020-3038
20

R. Chappel, B. Schwerin, and K. K. Paliwal, “Phase distortion resulting in a just noticeable difference in the perceived quality of speech,” Speech Commun. 81, 138-147 (2016).

10.1016/j.specom.2016.04.005
Information
  • Publisher :The Acoustical Society of Korea
  • Publisher(Ko) :한국음향학회
  • Journal Title :The Journal of the Acoustical Society of Korea
  • Journal Title(Ko) :한국음향학회지
  • Volume : 45
  • No :5
  • Pages :572-581
  • Received Date : 2026-07-14
  • Revised Date : 2026-08-11
  • Accepted Date : 2026-08-24