40 - On the State of the Art of Evaluation in Neural Language Models, with Gabor Melis

Published: Nov. 7, 2017, 7:58 p.m.

Recent arxiv paper by G\xe1bor Melis, Chris Dyer, and Phil Blunsom.\n\nG\xe1bor comes on the podcast to tell us about his work. He performs a thorough comparison between vanilla LSTMs and recurrent highway networks on the language modeling task, showing that when both methods are given equal amounts of hyperparameter tuning, LSTMs perform better, in contrast to prior work claiming that recurrent highway networks perform better. We talk about parameter tuning, training variance, language model evaluation, and other related issues.\n\nhttps://www.semanticscholar.org/paper/On-the-State-of-the-Art-of-Evaluation-in-Neural-La-Melis-Dyer/2397ce306e5d7f3d0492276e357fb1833536b5d8