Author
Listed:
- Yu Chang Yeh
- Ming Chieh Shih
- Daniel De Backer
- Leo Anthony Celi
- Kay Choong See
- Tomoko Fujii
- Lowell Ling
- Wasineenart Mongkolpun
- Hsiang Wei Hu
- Hsuan Yu Chen
- Wei Cheng Chen
- Bernard P. Cholley
- Kean Khang Fong
- Ho Geol Ryu
- Sungwon Na
- Moritoki Egi
- Wing Sum Chan
- Kuan Fu Chen
- Rishikesan Kamaleswaran
- Yu Chen Chuang
- Chi Ju Yang
- Wei Ling Hsiao
- Sheng Ru Lai
- David Ku
- Ahsina Jahan
- Greg Martin
Abstract
Background Generative artificial intelligence (GenAI) is increasingly used for clinical decision support in critical care, yet standardized methods for evaluating GenAI content in intensive care settings are lacking. Existing metrics assess textual similarity but fail to capture clinical accuracy, reasoning quality, or urgency. Methods We developed and validated the IMPACT framework through a five-phase multinational panel consensus process. Reporting adhered to the ACCORD guideline. A steering committee of eight persons provided clinical and methodological oversight. Panelists were recruited through purposive sampling to ensure geographic and multidisciplinary representation. Content validity was assessed using the Content Validity Ratio (CVR) and Item-level Content Validity Index (I-CVI), with retention thresholds set at 70% agreement and I-CVI ≥0.80. Results A total of 58 panelists from 12 countries and regions participated, with 42 completing formal consensus voting. Participants included intensivists, physicians with AI research expertise, information technology specialists, and other critical care professionals. All six IMPACT domains exceeded validity thresholds (mean agreement 89.3%, CVR = 0.79, I-CVI = 0.92). Of 24 candidate subitems, 21 met retention criteria (mean agreement 85.7%, CVR = 0.71, I-CVI = 0.90). Three subitems were removed due to insufficient consensus and conceptual overlap. The validated framework comprises six domains with 21 subitems. Conclusions The IMPACT framework provides a consensus-validated approach for evaluating GenAI clinical decision support in intensive care, addressing gaps in current evaluation methods.
Suggested Citation
Yu Chang Yeh & Ming Chieh Shih & Daniel De Backer & Leo Anthony Celi & Kay Choong See & Tomoko Fujii & Lowell Ling & Wasineenart Mongkolpun & Hsiang Wei Hu & Hsuan Yu Chen & Wei Cheng Chen & Bernard P, 2026.
"The IMPACT framework for evaluating generative AI in critical care: development and multinational consensus validation,"
ULB Institutional Repository
2013/413053, ULB -- Universite Libre de Bruxelles.
Handle:
RePEc:ulb:ulbeco:2013/413053
Note: SCOPUS: ar.j
Download full text from publisher
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:ulb:ulbeco:2013/413053. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
We have no bibliographic references for this item. You can help adding them by using this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Benoit Pauwels (email available below). General contact details of provider: https://edirc.repec.org/data/ecsulbe.html .
Please note that corrections may take a couple of weeks to filter through
the various RePEc services.