Barriers and enabling factors for error analysis in NLG research

Emiel van Miltenburg; Miruna Clinciu; Ondřej Dušek; Dimitra Gkatzia; Stephanie Inglis; Leo Leppänen; Saad Mahamood; Stephanie Schoch; Craig Thomson; Luou Wen

doi:10.3384/nejlt.2000-1533.2023.4529

Authors

Emiel van Miltenburg Tilburg University https://orcid.org/0000-0002-7143-8961
Miruna Clinciu Heriot-Watt University and the University of Edinburgh https://orcid.org/0000-0003-2233-3001
Ondřej Dušek Charles University, Prague https://orcid.org/0000-0002-1415-1702
Dimitra Gkatzia Edinburgh~Napier~University, UK https://orcid.org/0000-0001-8568-7806
Stephanie Inglis Arria NLG / University of Aberdeen https://orcid.org/0000-0001-7414-6086
Leo Leppänen University of Helsinki https://orcid.org/0000-0003-3969-8410
Saad Mahamood trivago N.V. https://orcid.org/0000-0003-2332-8749
Stephanie Schoch University of Virginia, USA https://orcid.org/0000-0003-2387-2189
Craig Thomson University of Aberdeen https://orcid.org/0000-0002-1602-4694
Luou Wen Independent researcher https://orcid.org/0000-0003-3296-3226

DOI:

https://doi.org/10.3384/nejlt.2000-1533.2023.4529

Abstract

Earlier research has shown that few studies in Natural Language Generation (NLG) evaluate their system outputs using an error analysis, despite known limitations of automatic evaluation metrics and human ratings. This position paper takes the stance that error analyses should be encouraged, and discusses several ways to do so. This paper is based on our shared experience as authors as well as a survey we distributed as a means of public consultation. We provide an overview of existing barriers to carrying out error analyses, and propose changes to improve error reporting in the NLG literature.

Barriers and enabling factors for error analysis in NLG research

Authors

DOI:

Abstract

Downloads

Published

Issue

Section

License

Make a Submission