Expediting data extraction using a large language model (LLM) and scoping review protocol: a methodological study within a complex scoping review

๐Ÿ“… 2025-07-09
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF

career value

159K/year
๐Ÿค– AI Summary
This study addresses the time- and resource-intensive bottleneck of data extraction in scoping reviews by proposing an LLM-driven automation framework leveraging Claude 3.5 Sonnet and a structured review protocol. Methodologically, the protocol is directly embedded into system prompts, and an LLM-in-the-loop feedback mechanism is introduced to iteratively refine protocol design. Results show high accuracy (83.3%โ€“100%) for simple, structured fields (e.g., author, year), but sharp declines for complex, subjective constructs (e.g., study aims, limitations)โ€”with accuracy dropping to 9.6%โ€“15.8% and overall F1 < 40%, primarily due to low recall (<25%). Protocol embedding improves inter-extraction consistency, and LLM feedback supports effective protocol refinement; however, current models exhibit insufficient reliability for semantic abstraction tasks. The study underscores the necessity of multi-dimensional performance evaluation and human-in-the-loop validation, offering both a methodological framework and pragmatic caution for deploying LLMs in evidence synthesis.

Technology Category

Application Category

๐Ÿ“ Abstract
The data extraction stages of reviews are resource-intensive, and researchers may seek to expediate data extraction using online (large language models) LLMs and review protocols. Claude 3.5 Sonnet was used to trial two approaches that used a review protocol to prompt data extraction from 10 evidence sources included in a case study scoping review. A protocol-based approach was also used to review extracted data. Limited performance evaluation was undertaken which found high accuracy for the two extraction approaches (83.3% and 100%) when extracting simple, well-defined citation details; accuracy was lower (9.6% and 15.8%) when extracting more complex, subjective data items. Considering all data items, both approaches had precision >90% but low recall (<25%) and F1 scores (<40%). The context of a complex scoping review, open response types and methodological approach likely impacted performance due to missed and misattributed data. LLM feedback considered the baseline extraction accurate and suggested minor amendments: four of 15 (26.7%) to citation details and 8 of 38 (21.1%) to key findings data items were considered to potentially add value. However, when repeating the process with a dataset featuring deliberate errors, only 2 of 39 (5%) errors were detected. Review-protocol-based methods used for expediency require more robust performance evaluation across a range of LLMs and review contexts with comparison to conventional prompt engineering approaches. We recommend researchers evaluate and report LLM performance if using them similarly to conduct data extraction or review extracted data. LLM feedback contributed to protocol adaptation and may assist future review protocol drafting.
Problem

Research questions and friction points this paper is trying to address.

Expediting data extraction using LLMs in scoping reviews
Evaluating accuracy of LLMs for simple vs complex data extraction
Assessing protocol-based methods for reviewing extracted data
Innovation

Methods, ideas, or system contributions that make the work stand out.

Using LLM for data extraction with review protocol
High accuracy for simple citation details extraction
Low recall for complex subjective data extraction
J
James Stewart-Evans
Nottingham Centre for Public Health and Epidemiology, University of Nottingham, Nottingham, UK; Environmental Hazards and Emergencies Department, UK Health Security Agency, UK
E
Emma Wilson
Nottingham Centre for Public Health and Epidemiology, University of Nottingham, Nottingham, UK; Centre for Evidence Based Healthcare*, University of Nottingham, Nottingham, UK
T
Tessa Langley
Nottingham Centre for Public Health and Epidemiology, University of Nottingham, Nottingham, UK
A
Andrew Prayle
Lifespan and Population Health, University of Nottingham, Nottingham, UK; Nottingham Biomedical Research Centre, University of Nottingham, Nottingham, UK
A
Angela Hands
Office for Health Improvement and Disparities, Department of Health and Social Care, UK
K
Karen Exley
Environmental Hazards and Emergencies Department, UK Health Security Agency, UK
J
Jo Leonardi-Bee
Nottingham Centre for Public Health and Epidemiology, University of Nottingham, Nottingham, UK; Centre for Evidence Based Healthcare*, University of Nottingham, Nottingham, UK