Multi-view Frequency LSTM: An Efficient Frontend for Automatic Speech Recognition

30 Jun 2020Maarten Van SegbroeckHarish MallidihBrian KingI-Fan ChenGurpreet ChadhaRoland Maas

Acoustic models in real-time speech recognition systems typically stack multiple unidirectional LSTM layers to process the acoustic frames over time. Performance improvements over vanilla LSTM architectures have been reported by prepending a stack of frequency-LSTM (FLSTM) layers to the time LSTM... (read more)

PDF Abstract

Code


No code implementations yet. Submit your code now

Results from the Paper


  Submit results from this paper to get state-of-the-art GitHub badges and help the community compare results to other papers.

Methods used in the Paper