Search Results for author: Kaustubh Kalgaonkar

Found 4 papers, 1 papers with code

Egocentric Audio-Visual Noise Suppression

no code implementations • 7 Nov 2022 • Roshan Sharma, Weipeng He, Ju Lin, Egor Lakomkin, Yang Liu, Kaustubh Kalgaonkar

In this paper, we first demonstrate that egocentric visual information is helpful for noise suppression.

Action Classification Event Detection +3

Paper
Add Code

SCA: Streaming Cross-attention Alignment for Echo Cancellation

no code implementations • 1 Nov 2022 • Yang Liu, Yangyang Shi, Yun Li, Kaustubh Kalgaonkar, Sriram Srinivasan, Xin Lei

End-to-End deep learning has shown promising results for speech enhancement tasks, such as noise suppression, dereverberation, and speech separation.

Speech Enhancement Speech Separation

Paper
Add Code

RNN-T For Latency Controlled ASR With Improved Beam Search

no code implementations • 5 Nov 2019 • Mahaveer Jain, Kjell Schubert, Jay Mahadeokar, Ching-Feng Yeh, Kaustubh Kalgaonkar, Anuroop Sriram, Christian Fuegen, Michael L. Seltzer

Neural transducer-based systems such as RNN Transducers (RNN-T) for automatic speech recognition (ASR) blend the individual components of a traditional hybrid ASR systems (acoustic model, language model, punctuation model, inverse text normalization) into one single model.

Automatic Speech Recognition Automatic Speech Recognition (ASR) +3

Paper
Add Code

Transformer-Transducer: End-to-End Speech Recognition with Self-Attention

1 code implementation • 28 Oct 2019 • Ching-Feng Yeh, Jay Mahadeokar, Kaustubh Kalgaonkar, Yongqiang Wang, Duc Le, Mahaveer Jain, Kjell Schubert, Christian Fuegen, Michael L. Seltzer

We explore options to use Transformer networks in neural transducer for end-to-end speech recognition.

speech-recognition Speech Recognition

Paper
Code

Cannot find the paper you are looking for? You can Submit a new open access paper.