Catalogue search • Linguistik portal • Fachinformationsdienst (FID)

1	End-to-end style-conditioned poetry generation: What does it take to learn from examples alone? ...
	The 2021 Conference on Empirical Methods in Natural Language Processing 2021; Eger, Steffen. - : Underline Science Inc., 2021
	BASE
	Show details

2	Global Explainability of BERT-Based Evaluation Metrics by Disentangling along Linguistic Factors ...
	The 2021 Conference on Empirical Methods in Natural Language Processing 2021; Eger, Steffen; Kaster, Marvin; Zhao, Wei. - : Underline Science Inc., 2021
	Abstract: Anthology paper link: https://aclanthology.org/2021.emnlp-main.701/ Abstract: Evaluation metrics are a key ingredient for progress of text generation systems. In recent years, several BERT-based evaluation metrics have been proposed (including BERTScore, MoverScore, BLEURT, etc.) which correlate much better with human assessment of text generation quality than BLEU or ROUGE, invented two decades ago. However, little is known what these metrics, which are based on black-box language model representations, actually capture (it is typically assumed they model semantic similarity). In this work, we use a simple regression based global explainability technique to disentangle metric scores along linguistic factors, including semantics, syntax, morphology, and lexical overlap. We show that the different metrics capture all aspects to some degree, but that they are all substantially sensitive to lexical overlap, just like BLEU and ROUGE. This exposes limitations of these novelly proposed metrics, which we also ...
	Keyword: Computational Linguistics; Language Models; Machine Learning; Machine Learning and Data Mining; Natural Language Processing
	URL: https://dx.doi.org/10.48448/aajb-9k90 https://underline.io/lecture/37492-global-explainability-of-bert-based-evaluation-metrics-by-disentangling-along-linguistic-factors
	BASE
	Hide details

Search in the Catalogues and Directories