Towards Neural Language Evaluators

Kané, Hassan, Kocyigit, Yusuf, Ajanoh, Pelkins, Abdalla, Ali, Coulibali, Mohamed

arXiv.org Artificial Intelligence 

W e review three limitations of BLEU and ROUGE - the most popul ar metrics used to assess reference summaries against hypothesis summ aries, come up with criteria for what a good metric should behave like and propos e concrete ways to use recent Transformers-based Language Models to assess re ference summaries against hypothesis summaries.

Duplicate Docs Excel Report

Title
None found

Similar Docs  Excel Report  more

TitleSimilaritySource
None found