arXiv · 1909.09268
Towards Neural Language Evaluators
Abstract
We review three limitations of BLEU and ROUGE -- the most popular metrics used to assess reference summaries against hypothesis summaries, come up with criteria for what a good metric should behave like and propose concrete ways to use recent Transformers-based Language Models to assess reference summaries against hypothesis summaries.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hassan Kané, Yusuf Kocyigit, Pelkins Ajanoh, Ali Abdalla, Mohamed Coulibali. 2019-10-30. Towards Neural Language Evaluators. https://arxiv.org/abs/1909.09268
Cite the original work for its findings. Save a collection to share your selection of sources.