Spectral Methods for Single Channel Speech Enhancement in Multi-Source Environment

Clicks: 3
ID: 313253
2022
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This article has not been analysed, so there is no overall score — reader engagement is measured and shown alongside.
AI Quality Assessment
Not analyzed
Readership in this journal
Emerging

Ranked #384 of 705 articles by views in Journal of Computing & Biomedical Informatics

Most read Least read

Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 705 in total.

Mint this article as an NFT
Not yet minted

Create a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.

5 SUSD one-off · no wallet required
Abstract
Speech communication for both humans and automatic devices can be negatively impacted by background noise, which is common in real environments. Among many techniques, speech separation using a single microphone is the most desirable from an application standpoint. The resulting monaural speech separation problem has been a central one in speech processing for several decades. However, its success has been limited thus far. This research presents work that develops speech separation systems using combinations of T-F masking, DNNs, and model-based reconstruction. The aim of each system is to improve the perceptual quality of the speech estimates. The performance of many speech processing applications is severely degraded when both noise and reverberation are present. The proposed solution has been tested in the simulation environment and based on the simulation result, it is observed that the speech enhancement can easily be performed through the integration of the solution. This research suggests two staged noise reducing systems in order to reduce the background noise through a single microphone recording in a low-SNR based on ideal binary masking and Wiener filter. It has two stages. Firstly, for background noise reduction, a Wiener filter with an enhanced SNR is utilised on noisy speech. Secondly, IBM is calculated in each time–frequency channel through utilisation of the pre–processed speech from the first stage and the matching of the time–frequency channels to a pre-selected threshold in order to minimise residual noise. These channels meeting the threshold requirement are conserved while all the other ones are attenuated.
Reference Key
imported_1777060049_69ebc8d169439 Use this key to autocite in the manuscript while using SciMatic Manuscript Manager or Thesis Manager
Authors Sheeraz Ahmad
Journal Journal of Computing & Biomedical Informatics
Year 2022
DOI
10.56979/302/2022/50
URL
Keywords Keywords not found

Citations

No citations found. To add a citation, contact the admin at info@scimatic.org

No comments yet. Be the first to comment on this article.