Vollständiger Abstract
Worum geht es in dieser Arbeit?
In practical voice communication and processing, speech signals are extremely sensitive to background noise, equipment noise, and interference from complex environments, leading to a decline in speech quality and clarity. This, in turn, affects the overall performance of downstream systems such as speech recognition, speech coding, and human-computer interaction. Therefore, researching efficient and robust single-channel speech signal enhancement methods has significant theoretical and practical implications. With the rapid development of deep learning technology, deep neural networks, with their superior nonlinear feature modeling capabilities and end-to-end training advantages, have become a core research area for single-channel speech enhancement. Given the limitations of traditional speech enhancement methods in complex, non-stationary, and noisy environments, this paper proposes a deep neural network-based speech enhancement method to address the challenges of single-channel speech enhancement. This method aims to extract a clearer, more natural, and more intelligible target signal from noisy speech. This paper first analyzes the basic principles of speech quality enhancement, including the characteristics of speech signals, types of background noise, time-domain and frequency-domain speech representation methods, and commonly used speech quality evaluation metrics. Based on this, this paper develops a deep neural network model for single-channel speech enhancement. This model effectively estimates the spectral information of the target speech by learning the spectral features of noisy speech.
Bibliografischer Nachweis
Publikationsdaten
- Autor:innen
- Wen Fan, Wei-Yu Liang, Duo-Duo Han, Wei Huang, Han-Jiao Meng
- Quelle
- 電腦學刊
- Publikation
- 2026-01-01
- Band / Ausgabe
- Nicht angegeben
- Seiten
- Nicht angegeben
- ISSN / ISBN
- 1991-1599, 2312-993X
- Zitationen
- 0 laut Crossref
- Referenzen
- 0 hinterlegt
Zitieren
Zitierfähiger Nachweis
Wen Fan, Wei-Yu Liang, Duo-Duo Han, Wei Huang, Han-Jiao Meng (2026). Single-Channel Speech Enhancement Method Based on Deep Neural Networks. 電腦學刊. https://doi.org/10.63367/199115992026083704017