SweClinEval: A Benchmark for Swedish Clinical Natural Language Processing

Vakili, Thomas; Hansson, Martin; Henriksson, Aron

SweClinEval: A Benchmark for Swedish Clinical Natural Language Processing

Failid

2025_nodalida_1_76.pdf (118.38 KB)

Kuupäev

2025-03

Autorid

Vakili, Thomas

Hansson, Martin

Henriksson, Aron

Kirjastaja

University of Tartu Library

Abstrakt

The lack of benchmarks in certain domains and for certain languages makes it difficult to track progress regarding the state-of-the-art of NLP in those areas, potentially impeding progress in important, specialized domains. Here, we introduce the first Swedish benchmark for clinical NLP: _SweClinEval_. The first iteration of the benchmark consists of six clinical NLP tasks, encompassing both document-level classification and named entity recognition tasks, with real clinical data. We evaluate nine different encoder models, both Swedish and multilingual. The results show that domain-adapted models outperform generic models on sequence-level classification tasks, while certain larger generic models outperform the clinical models on named entity recognition tasks. We describe how the benchmark can be managed despite limited possibilities to share sensitive clinical data, and discuss plans for extending the benchmark in future iterations.

URI

https://hdl.handle.net/10062/107269

Kollektsioonid

Proceedings of the Joint 25th Nordic Conference on Computational Linguistics and 11th Baltic Conference on Human Language Technologies (NoDaLiDa/Baltic-HLT 2025)

Kirje täielik lehekülg

SweClinEval: A Benchmark for Swedish Clinical Natural Language Processing

Failid

Kuupäev

Autorid

Ajakirja pealkiri

Ajakirja ISSN

Köite pealkiri

Kirjastaja

Abstrakt

Kirjeldus

Märksõnad

Viide

URI

Kollektsioonid