Characterizing the hyper-parameter space of LSTM language models for mixed context applications

Abstract

Applying state of the art deep learning models to novel real world datasetsgives a practical evaluation of the generalizability of these models. Ofimportance in this process is how sensitive the hyper parameters of such modelsare to novel datasets as this would affect the reproducibility of a model. Wepresent work to characterize the hyper parameter space of an LSTM for languagemodeling on a code-mixed corpus. We observe that the evaluated model showsminimal sensitivity to our novel dataset bar a few hyper parameters.

Quick Read (beta)

loading the full paper ...