Underreporting of errors in NLG output, and what to do about it

Emiel van Miltenburg,Miruna-Adriana Clinciu,Ondřej Dušek,Dimitra Gkatzia,Stephanie Inglis,Leo Leppänen,Saad Mahamood,Emma Manning,Stephanie Schoch,Craig Thomson,Luou Wen

Underreporting of errors in NLG output, and what to do about it

2021

Emiel van Miltenburg
Miruna-Adriana Clinciu
Ondřej Dušek
Dimitra Gkatzia
Stephanie Inglis
Leo Leppänen
Saad Mahamood
Emma Manning
Stephanie Schoch
Craig Thomson
Luou Wen

We observe a severe under-reporting of the different kinds of errors that Natural Language Generation systems make. This is a problem, because mistakes are an important indicator of where systems should still be improved. If authors only report overall performance metrics, the research community is left in the dark about the specific weaknesses that are exhibited by `state-of-the-art' research. Next to quantifying the extent of error under-reporting, this position paper provides recommendations for error identification, analysis and reporting.

Keywords:

Artificial intelligence
Machine learning
Computer science
error identification
Natural language generation
research community
overall performance
Position paper

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations