Associative Memory Model-Based Linear Filtering and Its Application to Tandem Connectionist Blind Source Separation

    研究成果: Article

    2 引用 (Scopus)

    抜粋

    We propose a blind source separation method that yields high-quality speech with low distortion. Time-frequency (TF) masking can effectively reduce interference, but it produces nonlinear distortion. By contrast, linear filtering using a separation matrix such as independent vector analysis (IVA) can avoid nonlinear distortion, but the separation performance is reduced under reverberant conditions. The tandem connectionist approach combines several separation methods and it has been used frequently to compensate for the disadvantages of these methods. In this study, we propose associative memory model (AMM)-based linear filtering and a tandem connectionist framework, which applies TF masking followed by linear filtering. By using AMM trained with speech spectra to optimize the separation matrix, the proposed linear filtering method considers the properties of speech that are not considered explicitly in IVA, such as the harmonic components of spectra. TF masking is applied in the proposed tandem connectionist framework to reduce unwanted components that hinder the optimization of the separation matrix, and it is approximated by using a linear separation matrix to reduce nonlinear distortion. The results obtained in simultaneous speech separation experiments demonstrate that although the proposed linear filtering method can increase the signal-to-distortion ratio (SDR) and signal-to-interference ratio (SIR) compared with IVA, the proposed tandem connectionist framework can obtain greater increases in SDR and SIR, and it reduces the phoneme error rate more than the proposed linear filtering method.

    元の言語English
    記事番号7819470
    ページ(範囲)637-650
    ページ数14
    ジャーナルIEEE/ACM Transactions on Audio Speech and Language Processing
    25
    発行部数3
    DOI
    出版物ステータスPublished - 2017 3 1

      フィンガープリント

    ASJC Scopus subject areas

    • Signal Processing
    • Media Technology
    • Instrumentation
    • Acoustics and Ultrasonics
    • Linguistics and Language
    • Speech and Hearing
    • Electrical and Electronic Engineering

    これを引用