Author
Listed:
- Luo, Ronald
- Faisal, Abu Ilius
- Sastimoglu, Ziya
- Deen, Jamal
Abstract
Background: Systematic reviews are crucial for evidence-based practice but are often hindered by labor-intensive methodologies. Recent advances in Large Language Models (LLMs) may help improve review efficiency and replicability. Objective: Building on prior research, this study evaluates the impact of prompt engineering on LLM performance in automating systematic reviews, specifically examining the "zero-shot" capability for predicting article inclusion based on titles and abstracts alone. Methods: We reanalyzed 24,534 studies across three systematic reviews using two sets of prompts with GPT-3.5 Turbo. One set rated the suitability of a study on a scale of one to ten, while the other elicited a binary response. Both ordinal-response and binary-response prompts were assessed for accuracy, sensitivity, specificity, predictive values, F1 score, and Matthews Correlation Coefficient (MCC). We also evaluated Receiver Operating Characteristic (ROC) curves for ordinal-response prompts and statistically validated the results using K-fold permutations and one-way analysis of variance. Results: Ordinal prompting allowed for fine-tuning of sensitivity and specificity, proving to be a viable zero-shot inclusion prompt with ROC Area Under Curve (AUC) scores between 80-90%. Including studies rated seven or higher achieved sensitivity and specificity of ~80%, while using the mean score increased sensitivity to nearly 100%, a task not possible with earlier strategies. Conclusion: Prompt engineering offers a scalable solution that allows researchers to target highly relevant articles or maximize sensitivity to include nearly all relevant studies if more time is available.
Suggested Citation
Luo, Ronald & Faisal, Abu Ilius & Sastimoglu, Ziya & Deen, Jamal, 2026.
"Ordinal Response Elicitation for Systematic Review Screening with GPT,"
MetaArXiv
tvekx_v1, Center for Open Science.
Handle:
RePEc:osf:metaar:tvekx_v1
DOI: 10.31219/osf.io/tvekx_v1
Download full text from publisher
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:osf:metaar:tvekx_v1. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
We have no bibliographic references for this item. You can help adding them by using this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: OSF (email available below). General contact details of provider: https://osf.io/preprints/metaarxiv .
Please note that corrections may take a couple of weeks to filter through
the various RePEc services.