DE eng

Search in the Catalogues and Directories

Hits 1 – 5 of 5

1
Better than Average: Paired Evaluation of NLP systems ...
BASE
Show details
2
Inducing Language-Agnostic Multilingual Representations ...
BASE
Show details
3
Global Explainability of BERT-Based Evaluation Metrics by Disentangling along Linguistic Factors ...
Abstract: Evaluation metrics are a key ingredient for progress of text generation systems. In recent years, several BERT-based evaluation metrics have been proposed (including BERTScore, MoverScore, BLEURT, etc.) which correlate much better with human assessment of text generation quality than BLEU or ROUGE, invented two decades ago. However, little is known what these metrics, which are based on black-box language model representations, actually capture (it is typically assumed they model semantic similarity). In this work, we use a simple regression based global explainability technique to disentangle metric scores along linguistic factors, including semantics, syntax, morphology, and lexical overlap. We show that the different metrics capture all aspects to some degree, but that they are all substantially sensitive to lexical overlap, just like BLEU and ROUGE. This exposes limitations of these novelly proposed metrics, which we also highlight in an adversarial test scenario. ... : EMNLP2021 Camera Ready ...
Keyword: Computation and Language cs.CL; FOS Computer and information sciences
URL: https://arxiv.org/abs/2110.04399
https://dx.doi.org/10.48550/arxiv.2110.04399
BASE
Hide details
4
Global Explainability of BERT-Based Evaluation Metrics by Disentangling along Linguistic Factors ...
BASE
Show details
5
The Trans-Ancestral Genomic Architecture of Glycemic Traits
In: Nat Genet (2021)
BASE
Show details

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
5
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern