Catalogue search • Linguistik portal • Fachinformationsdienst (FID)

1	IAPUCP at SemEval-2021 task 1: Stacking fine-tuned transformers is almost all you need for lexical complexity prediction
	Rivas Rojas, Kervy; Alva-Manchego, Fernando. - : Association for Computational Linguistics, 2021
	Abstract: This paper describes our submission to SemEval-2021 Task 1: predicting the complexity score for single words. Our model leverages standard morphosyntactic and frequency-based features that proved helpful for Complex Word Identification (a related task), and combines them with predictions made by Transformer-based pre-trained models that were fine-tuned on the Shared Task data. Our submission system stacks all previous models with a LightGBM at the top. One novelty of our approach is the use of multi-task learning for fine-tuning a pre-trained model for both Lexical Complexity Prediction and Word Sense Disambiguation. Our analysis shows that all independent models achieve a good performance in the task, but that stacking them obtains a Pearson correlation of 0.7704, merely 0.018 points behind the winning submission.
	URL: https://doi.org/10.18653/v1/2021.semeval-1.14 https://orca.cardiff.ac.uk/147258/1/2021.semeval-1.14.pdf https://orca.cardiff.ac.uk/147258/
	BASE
	Hide details

Search in the Catalogues and Directories