Archive \ Volume.17 2026 Issue 3

Why Pharmacy AI Evaluation Must Begin with Plausible Harm Scenarios

, , ,
  1. Department of AI Harm Scenario Evaluation, Faculty of Pharmacy, University of Bucharest, Bucharest, Romania.
  2. Department of Pharmacy AI Safety Assessment, Faculty of Pharmacy, University of Agricultural Sciences Cluj-Napoca, Cluj-Napoca, Romania.
  3. Department of Plausible Harm Analysis in Pharmacy, Faculty of Pharmacy, Polytechnic University of Bucharest, Bucharest, Romania.

Abstract

Artificial intelligence evaluation in pharmacy commonly begins with model-centred measures such as discrimination, sensitivity, specificity, calibration, alert acceptance, or processing efficiency. These measures are necessary, but aggregate performance can obscure how errors affect particular patients, professionals, medication-use tasks, and organizational conditions. This article develops an original, non-empirical Plausible-Harm Scenario Framework for Pharmacy AI Evaluation. The proposed approach begins by specifying who may be affected, which medication-use task is involved, how an AI-related failure could propagate through data, model output, interface presentation, professional interpretation, and action or inaction, and what medication-related consequence could follow. Each scenario is then examined through five distinct dimensions: severity, exposure, detectability, reversibility, and recovery. These dimensions are not combined into a numerical score. Instead, they organize the selection of technical, clinical-task, human–AI, workflow, medication-safety, equity, implementation, and lifecycle evidence. The framework further proposes governance gates for scenario completeness, evidence adequacy, residual-risk deliberation, bounded permission, monitoring, rollback, and requalification. Its principal contribution is to reposition performance metrics as scenario-dependent evidence rather than sufficient indicators of safety or readiness. Validation would require prospective and post-deployment testing across relevant users, settings, populations, workflows, model versions, and failure conditions. The framework cannot guarantee complete hazard discovery, establish universal thresholds, replace professional judgment, or demonstrate clinical benefit. It is intended as a conceptual structure for making the consequences and evidentiary assumptions of pharmacy AI evaluation more explicit.


Downloads: 20
Views: 81

How to cite:
Vancouver
Popescu A, Ionescu M, Stan E, Radu C. Why Pharmacy AI Evaluation Must Begin with Plausible Harm Scenarios. Arch Pharm Pract. 2026;17(3):28-35. https://doi.org/10.51847/7eukyIeyrZ
APA
Popescu, A., Ionescu, M., Stan, E., & Radu, C. (2026). Why Pharmacy AI Evaluation Must Begin with Plausible Harm Scenarios. Archives of Pharmacy Practice, 17(3), 28-35. https://doi.org/10.51847/7eukyIeyrZ

Download Citation
References
  1. Kelly CJ, Karthikesalingam A, Suleyman M, Corrado G, King D. Key challenges for delivering clinical impact with artificial intelligence. BMC Med. 2019;17(1):195. doi:10.1186/s12916-019-1426-2
  2. Liu X, Faes L, Kale AU, Wagner SK, Fu DJ, Bruynseels A, et al. A comparison of deep learning performance against health-care professionals in detecting diseases from medical imaging: A systematic review and meta-analysis. Lancet Digit Health. 2019;1(6):e271-e97. doi:10.1016/S2589-7500(19)30123-2
  3. Wiens J, Saria S, Sendak M, Ghassemi M, Liu VX, Doshi-Velez F, et al. Do no harm: A roadmap for responsible machine learning for health care. Nat Med. 2019;25(9):1337-40. doi:10.1038/s41591-019-0548-6
  4. Cabitza F, Rasoini R, Gensini GF. Unintended consequences of machine learning in medicine. JAMA. 2017;318(6):517-8. doi:10.1001/jama.2017.7797
  5. Nagendran M, Chen Y, Lovejoy CA, Gordon AC, Komorowski M, Harvey H, et al. Artificial intelligence versus clinicians: Systematic review of design, reporting standards, and claims of deep learning studies. BMJ. 2020;368:m689. doi:10.1136/bmj.m689
  6. Vasey B, Nagendran M, Campbell B, Clifton DA, Collins GS, Denaxas S, et al. Reporting guideline for the early-stage clinical evaluation of decision support systems driven by artificial intelligence: DECIDE-AI. Nat Med. 2022;28(5):924-33. doi:10.1038/s41591-022-01772-9
  7. Hernandez-Boussard T, Bozkurt S, Ioannidis JPA, Shah NH. MINIMAR (MINimum Information for Medical AI Reporting): Developing reporting standards for artificial intelligence in health care. J Am Med Inform Assoc. 2020;27(12):2011-5. doi:10.1093/jamia/ocaa088
  8. Van Calster B, McLernon DJ, van Smeden M, Wynants L, Steyerberg EW. Calibration: The Achilles heel of predictive analytics. BMC Med. 2019;17(1):230. doi:10.1186/s12916-019-1466-7
  9. Sutton RT, Pincock D, Baumgart DC, Sadowski DC, Fedorak RN, Kroeker KI. An overview of clinical decision support systems: Benefits, risks, and strategies for success. NPJ Digit Med. 2020;3:17. doi:10.1038/s41746-020-0221-y
  10. Reddy S, Allan S, Coghlan S, Cooper P. A governance model for the application of AI in health care. J Am Med Inform Assoc. 2020;27(3):491-7. doi:10.1093/jamia/ocz192
  11. Bates DW, Levine D, Syrowatka A, Kuznetsova M, Craig KJT, Rui A, et al. The potential of artificial intelligence to improve patient safety: A scoping review. NPJ Digit Med. 2021;4(1):54. doi:10.1038/s41746-021-00423-6
  12. Tschandl P, Rinner C, Apalla Z, Argenziano G, Codella N, Halpern A, et al. Human-computer collaboration for skin cancer recognition. Nat Med. 2020;26(8):1229-34. doi:10.1038/s41591-020-0942-0
  13. Kiani A, Uyumazturk B, Rajpurkar P, Wang A, Gao R, Jones E, et al. Impact of a deep learning assistant on the histopathologic classification of liver cancer. NPJ Digit Med. 2020;3:23. doi:10.1038/s41746-020-0232-8
  14. Gaube S, Suresh H, Raue M, Merritt A, Berkowitz SJ, Lermer E, et al. Do as AI say: Susceptibility in deployment of clinical decision-aids. NPJ Digit Med. 2021;4(1):31. doi:10.1038/s41746-021-00385-9
  15. Graafsma J, Murphy RM, van de Garde EMW, Karapinar-Çarkit F, Derijks HJ, Hoge RHL, et al. The use of artificial intelligence to optimize medication alerts generated by clinical decision support systems: A scoping review. J Am Med Inform Assoc. 2024;31(6):1411-22. doi:10.1093/jamia/ocae076
  16. Levivien C, Cavagna P, Grah A, Buronfosse A, Courseau R, Bézie Y, et al. Assessment of a hybrid decision support system using machine learning with artificial intelligence to safely rule out prescriptions from medication review in daily practice. Int J Clin Pharm. 2022;44(2):459-65. doi:10.1007/s11096-021-01366-4
  17. Chen CY, Chen YL, Scholl J, Yang HC, Li YCJ. Ability of machine-learning based clinical decision support system to reduce alert fatigue, wrong-drug errors, and alert users about look alike, sound alike medication. Comput Methods Programs Biomed. 2024;243:107869. doi:10.1016/j.cmpb.2023.107869
  18. Davis SE, Lasko TA, Chen G, Siew ED, Matheny ME. Calibration drift in regression and machine learning models for acute kidney injury. J Am Med Inform Assoc. 2017;24(6):1052-61. doi:10.1093/jamia/ocx030
  19. Subbaswamy A, Saria S. From development to deployment: Dataset shift, causality, and shift-stable models in health AI. Biostatistics. 2020;21(2):345-52. doi:10.1093/biostatistics/kxz041
  20. Begoli E, Bhattacharya T, Kusnezov D. The need for uncertainty quantification in machine-assisted medical decision making. Nat Mach Intell. 2019;1(1):20-3. doi:10.1038/s42256-018-0004-1
  21. Char DS, Shah NH, Magnus D. Implementing machine learning in health care—addressing ethical challenges. N Engl J Med. 2018;378(11):981-3. doi:10.1056/NEJMp1714229
  22. Gianfrancesco MA, Tamang S, Yazdany J, Schmajuk G. Potential biases in machine learning algorithms using electronic health record data. JAMA Intern Med. 2018;178(11):1544-7. doi:10.1001/jamainternmed.2018.3763
  23. Muehlematter UJ, Daniore P, Vokinger KN. Approval of artificial intelligence and machine learning-based medical devices in the USA and Europe (2015-20): A comparative analysis. Lancet Digit Health. 2021;3(3):e195-e203. doi:10.1016/S2589-7500(20)30292-2
  24. Futoma J, Simons M, Panch T, Doshi-Velez F, Celi LA. The myth of generalisability in clinical research and machine learning in health care. Lancet Digit Health. 2020;2(9):e489-e92. doi:10.1016/S2589-7500(20)30186-2
  25. Chen IY, Pierson E, Rose S, Joshi S, Ferryman K, Ghassemi M. Ethical machine learning in healthcare. Annu Rev Biomed Data Sci. 2021;4:123-44. doi:10.1146/annurev-biodatasci-092820-114757
  26. Norgeot B, Quer G, Beaulieu-Jones BK, Torkamani A, Dias R, Gianfrancesco M, et al. Minimum information about clinical artificial intelligence modeling: The MI-CLAIM checklist. Nat Med. 2020;26(9):1320-4. doi:10.1038/s41591-020-1041-y

 

 

 

 


Creative Commons License
This work is licensed under a Creative Commons Attribution 4.0 International License.