Abstract:Background: Radiology reports are typically written in a free-text format, making clinical information difficult to extract and use. Recently, the adoption of structured reporting (SR) has been recommended by various medical societies thanks to the advantages it offers, e.g. standardization, completeness, and information retrieval. We propose a pipeline to extract information from Italian free-text radiology reports that fits with the items of the reference SR registry proposed by a national society of interventional and medical radiology, focusing on CT staging of patients with lymphoma. Methods: Our work aims to leverage the potential of Natural Language Processing and Transformer-based models to deal with automatic SR registry filling. With the availability of 174 Italian radiology reports, we investigate a rule-free generative Question Answering approach based on the Italian-specific version of T5: IT5. To address information content discrepancies, we focus on the six most frequently filled items in the annotations made on the reports: three categorical (multichoice), one free-text (free-text), and two continuous numerical (factual). In the preprocessing phase, we encode also information that is not supposed to be entered. Two strategies (batch-truncation and ex-post combination) are implemented to comply with the IT5 context length limitations. Performance is evaluated in terms of strict accuracy, f1, and format accuracy, and compared with the widely used GPT-3.5 Large Language Model. Unlike multichoice and factual, free-text answers do not have 1-to-1 correspondence with their reference annotations. For this reason, we collect human-expert feedback on the similarity between medical annotations and generated free-text answers, using a 5-point Likert scale questionnaire (evaluating the criteria of correctness and completeness). Results: The combination of fine-tuning and batch splitting allows IT5 ex-post combination to achieve notable results in terms of information extraction of different types of structured data, performing on par with GPT-3.5. Human-based assessment scores of free-text answers show a high correlation with the AI performance metrics f1 (Spearman's correlation coefficients>0.5, p-values<0.001) for both IT5 ex-post combination and GPT-3.5. The latter is better at generating plausible human-like statements, even if it systematically provides answers even when they are not supposed to be given. Conclusions: In our experimental setting, a fine-tuned Transformer-based model with a modest number of parameters (i.e., IT5, 220 M) performs well as a clinical information extraction system for automatic SR registry filling task. It can extract information from more than one place in the report, elaborating it in a manner that complies with the response specifications provided by the SR registry (for multichoice and factual items), or that closely approximates the work of a human-expert (free-text items); with the ability to discern when an answer is supposed to be given or not to a user query.

Reshaping free-text radiology notes into structured reports with generative question answering transformers

Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers

An Inclusive Task-Aware Framework for Radiology Report Generation

Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports

Application of Deep Learning in Generating Structured Radiology Reports: A Transformer-Based Technique

Information extraction from weakly structured radiological reports with natural language queries

Transforming free-text radiology reports into structured reports using ChatGPT: A study on thyroid ultrasonography

Extracting Pulmonary Nodules and Nodule Characteristics from Radiology Reports of Lung Cancer Screening Patients Using Transformer Models

Large language models for structured reporting in radiology: past, present, and future

Toward an enhanced automatic medical report generator based on large transformer models

Automatic extraction of imaging observation and assessment categories from breast magnetic resonance imaging reports with natural language processing.

Automating Clinical Chart Review: An Open-Source Natural Language Processing Pipeline Developed on Free-Text Radiology Reports From Patients With Glioblastoma

Practical Evaluation of ChatGPT Performance for Radiology Report Generation

Natural Language Processing Technologies in Radiology Research and Clinical Applications.

Clinical Context-aware Radiology Report Generation from Medical Images using Transformers

Building blocks for complex tasks: Robust generative event extraction for radiology reports under domain shifts

Natural Language Processing Model for Identifying Critical Findings-A Multi-Institutional Study

Translating radiology reports into plain language using ChatGPT and GPT-4 with prompt learning: results, limitations, and potential

Natural Language Processing to extract SNOMED-CT codes from pathological reports

Using GPT‐4 for LI‐RADS feature extraction and categorization with multilingual free‐text reports

Learning Semi-Structured Representations of Radiology Reports