Information-Weighted Neural Cache Language Models for ASR.

Lyan Verwimp,Joris Pelemans,Hugo Van hamme,Patrick Wambacq

2018 IEEE Spoken Language Technology Workshop (SLT)（2018）

引用 24|浏览29

暂无评分

摘要

Neural cache language models (LMs) extend the idea of regular cache language models by making the cache probability dependent on the similarity between the current context and the context of the words in the cache. We make an extensive comparison of ‘regular’ cache models with neural cache models, both in terms of perplexity and WER after rescoring first-pass ASR results. Furthermore, we propose two extensions to this neural cache model that make use of the content value/information weight of the word: firstly, combining the cache probability and LM probability with an information-weighted interpolation and secondly, selectively adding only content words to the cache. We obtain a 29.9%/32.1% (validation/test set) relative improvement in perplexity with respect to a baseline LSTM LM on theWikiText-2 dataset, outperforming previous work on neural cache LMs. Additionally, we observe significant WER reductions with respect to the baseline model on the WSJ ASR task.

查看译文

关键词

Interpolation,Probability,Mathematical model,Context modeling,Speech recognition,Training data,Training

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要