ub.xmlui.mirage2.page-structure.muninLogoub.xmlui.mirage2.page-structure.openResearchArchiveLogo
    • EnglishEnglish
    • norsknorsk
  • Velg spraaknorsk 
    • EnglishEnglish
    • norsknorsk
  • Administrasjon/UB
Vis innførsel 
  •   Hjem
  • Fakultet for naturvitenskap og teknologi
  • Institutt for informatikk
  • Artikler, rapporter og annet (informatikk)
  • Vis innførsel
  •   Hjem
  • Fakultet for naturvitenskap og teknologi
  • Institutt for informatikk
  • Artikler, rapporter og annet (informatikk)
  • Vis innførsel
JavaScript is disabled for your browser. Some features of this site may not work without it.

Deidentifying a Norwegian clinical corpus - An effort to create a privacy-preserving Norwegian large clinical language model

Permanent lenke
https://hdl.handle.net/10037/33415
Thumbnail
Åpne
article.pdf (313.3Kb)
Akseptert manusversjon (PDF)
Dato
2024
Type
Journal article
Tidsskriftartikkel
Peer reviewed

Forfatter
Ngo, Phuong Dinh; Tejedor Hernandez, Miguel Angel; Olsen Svenning, Therese; Chomutare, Taridzo Fred; Budrionis, Andrius; Dalianis, Hercules
Sammendrag
This study discusses the methods and challenges of deidentifying and pseudonymizing Norwegian clinical text for research purposes. The results of the NorDeid tool for deidentification and pseudonymization on different types of protected health information were evaluated and discussed, as well as the extension of its functionality with regular expressions to identify specific types of sensitive information. This research used a clinical corpus of adult patients treated in a gastro-surgical department in Norway, which contains approximately nine million clinical notes. The study also highlights the challenges posed by the unique language and clinical terminology of Norway and emphasizes the importance of protecting privacy and the need for customized approaches to meet legal and research requirements.
Beskrivelse
Source at https://aclanthology.org/2024.caldpseudo-1.0.
Forlag
ACL
Sitering
Ngo, Tejedor Hernandez, Olsen Svenning, Chomutare, Budrionis, Dalianis. Deidentifying a Norwegian clinical corpus - An effort to create a privacy-preserving Norwegian large clinical language model. Proceedings of the Workshop on Computational Approaches to Language Data Pseudonymization (CALD-pseudo 2024). 2024
Metadata
Vis full innførsel
Samlinger
  • Artikler, rapporter og annet (informatikk) [478]
Copyright 2024 The Author(s)

Bla

Bla i hele MuninEnheter og samlingerForfatterlisteTittelDatoBla i denne samlingenForfatterlisteTittelDato
Logg inn

Statistikk

Antall visninger
UiT

Munin bygger på DSpace

UiT Norges Arktiske Universitet
Universitetsbiblioteket
uit.no/ub - munin@ub.uit.no

Tilgjengelighetserklæring