Generating Instructions in Virtual Environments (GIVE): A Challenge and an Evaluation Testbed for NLG

Donna K. Byron, Alexander Koller, Jon Oberlander, Laura Stoia, Kristina Striegnitz · 2007

Would it be helpful or detrimental for the field of NLG to have a generally accepted competition? Competitions have definitely advanced the state of the art in some fields of NLP, but the benefits sometimes come at the price of over-competitiveness, and there is a danger of overfitting systems to the concrete evaluation metrics. Moreover, it has been argued that there are intrinsic difficulties in NLG that make it harder to evaluate than other NLP tasks (Scott and Moore, 2006).

Read the paper · More papers on PaperTik