Proceedings chapter
OA Policy
English

Superlim: A Swedish Language Understanding Evaluation Benchmark

Presented atSingapore, 6 –10 December 2023
PublisherSingapore : Association for Computational Linguistics
Publication date2023-12-01
Abstract

We present Superlim, a multi-task NLP benchmark and analysis platform for evaluating Swedish language models, a counterpart to the English-language (Super)GLUE suite. We describe the dataset, the tasks, the leaderboard and report the baseline results yielded by a reference implementation. The tested models do not approach ceiling performance on any of the tasks, which suggests that Superlim is truly difficult, a desirable quality for a benchmark. We address methodological challenges, such as mitigating the Anglocentric bias when creating datasets for a less-resourced language; choosing the most appropriate measures; documenting the datasets and making the leaderboard convenient and transparent. We also highlight other potential usages of the dataset, such as, for instance, the evaluation of cross-lingual transfer learning.

Citation (ISO format)
BERDICEVSKIS, Aleksandrs et al. Superlim: A Swedish Language Understanding Evaluation Benchmark. In: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing. Singapore. Singapore : Association for Computational Linguistics, 2023. p. 8137–8153. doi: 10.18653/v1/2023.emnlp-main.506
Main files (1)
Proceedings chapter (Published version)
Identifiers
Additional URL for this publicationhttps://aclanthology.org/2023.emnlp-main.506/
78views
724downloads

Technical informations

Creation01/07/2024 06:09:03
First validation11/07/2024 07:25:12
Update30/08/2024 07:39:57
Status update30/08/2024 07:39:57
Last indexation01/11/2024 10:16:18
All rights reserved by Archive ouverte UNIGE and the University of GenevaunigeBlack