Skip to main navigation Skip to search Skip to main content

Rare Codes Count: Mining Inter-code Relations for Long-tail Clinical Text Classification

  • Jiamin Chen
  • , Xuhong Li
  • , Junting Xi
  • , Lei Yu
  • , Haoyi Xiong
  • Beihang University
  • Baidu Inc

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Multi-label clinical text classification, such as automatic ICD coding, has always been a challenging subject in Natural Language Processing, due to its long, domain-specific documents and long-tail distribution over a large label set. Existing methods adopt different model architectures to encode the clinical notes. Whereas without digging out the useful connections between labels, the model presents a huge gap in predicting performances between rare and frequent codes. In this work, we propose a novel method for further mining the helpful relations between different codes via a relationenhanced code encoder to improve the rare code performance. Starting from the simple code descriptions, the model reaches comparable, even better performances than models with heavy external knowledge. Our proposed method is evaluated on MIMIC-III, a common dataset in the medical domain. It outperforms the previous state-of-art models on both overall metrics and rare code performances. Moreover, the interpretation results further prove the effectiveness of our methods. Our code is publicly available1 .

Original languageEnglish
Title of host publication5th Workshop on Clinical Natural Language Processing, ClinicalNLP 2023 - Proceedings of the Workshop
PublisherAssociation for Computational Linguistics (ACL)
Pages403-413
Number of pages11
ISBN (Electronic)9781959429883
StatePublished - 2023
Event5th Workshop on Clinical Natural Language Processing, ClinicalNLP 2023. held at ACL 2023 - Toronto, Canada
Duration: 14 Jul 2023 → …

Publication series

NameProceedings of the Annual Meeting of the Association for Computational Linguistics
ISSN (Print)0736-587X

Conference

Conference5th Workshop on Clinical Natural Language Processing, ClinicalNLP 2023. held at ACL 2023
Country/TerritoryCanada
CityToronto
Period14/07/23 → …

Fingerprint

Dive into the research topics of 'Rare Codes Count: Mining Inter-code Relations for Long-tail Clinical Text Classification'. Together they form a unique fingerprint.

Cite this