Efficient Trainable Front-Ends for Neural Speech Enhancement

Casebeer, Jonah, Isik, Umut, Venkataramani, Shrikant, Krishnaswamy, Arvindh

Feb-19-2020–arXiv.org Machine Learning

Many neural speech enhancement and source separation systems operate in the time-frequency domain. Such models often benefit from making their Short-Time Fourier Transform (STFT) front-ends trainable. In current literature, these are implemented as large Discrete Fourier Transform matrices; which are prohibitively inefficient for low-compute systems. We present an efficient, trainable front-end based on the butterfly mechanism to compute the Fast Fourier Transform, and show its accuracy and efficiency benefits for low-compute neural speech enhancement models. We also explore the effects of making the STFT window trainable.

matrix, opération, speech enhancement, (11 more...)

arXiv.org Machine Learning

Feb-19-2020

arXiv.org PDF

Add feedback

Country:
- South America > Chile
  - Santiago Metropolitan Region > Santiago Province > Santiago (0.04)
- North America > United States
  - Illinois (0.04)

Genre:
- Research Report (0.50)

Technology:
- Information Technology
  - Data Science > Data Quality
    - Data Transformation (1.00)
  - Artificial Intelligence
    - Machine Learning (1.00)
    - Speech (0.94)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found