Towards Neural Language Evaluators
Kané, Hassan, Kocyigit, Yusuf, Ajanoh, Pelkins, Abdalla, Ali, Coulibali, Mohamed
–arXiv.org Artificial Intelligence
W e review three limitations of BLEU and ROUGE - the most popul ar metrics used to assess reference summaries against hypothesis summ aries, come up with criteria for what a good metric should behave like and propos e concrete ways to use recent Transformers-based Language Models to assess re ference summaries against hypothesis summaries.
arXiv.org Artificial Intelligence
Sep-19-2019
- Country:
- North America
- Canada (0.04)
- United States > Massachusetts
- Middlesex County > Cambridge (0.05)
- North America
- Genre:
- Research Report (0.40)
- Technology: