Accessibility settings

Published on in Vol 5 (2026)

Preprints (earlier versions) of this paper are available at https://preprints.jmir.org/preprint/90759, first published .
Laptop displaying a digital medical security network with a magnifying glass and shields.

Toward Retrieval-Grounded Evaluation for Conversational Large Language Model–Based Risk Assessment

Toward Retrieval-Grounded Evaluation for Conversational Large Language Model–Based Risk Assessment

Authors of this article:

Yihan Hu1 Author Orcid Image

Journals

  1. Roshani M, Zhou X, Qiang Y, Suresh S, Hicks S, Sethuraman U, Zhu D. Authors’ Reply: Toward Retrieval-Grounded Evaluation for Conversational Large Language Model–Based Risk Assessment. JMIR AI 2026;5:e91981 View
  2. Cai C, Zhu G, Zhou S, Ng O, Duell J, Ho W, Chen D, Lee B, Li F, Liu S, Sudarshan V, Wang L, Choi C, Fan X. A Competency Framework for Medical AI Education: Mixed Methods Study. JMIR Medical Education 2026;12:e91116 View
  3. Kim M, An Y, Jeon M, Lee Y, Lahcine O, Kim H, Lee S, Lee S, Yang J, Jeon S, Jung D, Cho C. Mapping Practice-Based Signals of Generative AI in Psychiatric Care: Qualitative Study of Korean Psychiatrists’ Experiences, Interpretations, and Implementation Priorities. Journal of Medical Internet Research 2026;28:e96556 View
  4. Kumar A, Varshney L. Training AI Models for Aesthetic Facial Evaluation: Focused Review and Framework to Mitigate Homogenizing Bias. Journal of Medical Internet Research 2026;28:e95452 View
  5. Qeyam H, Al-Rusan A. ChatGPT-Generated Advice on Sun Protection and Skin Cancer Prevention Compared to American Academy of Dermatology Guidelines: Cross-Sectional Content Analysis. JMIR Dermatology 2026;9:e93839 View
  6. Khatami A, Abbaspour-Raddakheli F, Wiefels C, Leung E. Exploring the Role of AI in Enhancing Nuclear Medicine Report Impressions Generated by Trainees and ChatGPT-4o: Comparative Evaluation Study. JMIR AI 2026;5:e94833 View
  7. Di Pumpo M, Villani L, Gualano M, Buonsenso D, Raffaelli F, Donà D, Laurenti P, Maio V, Boccia S, Ricciardi W. LLM-as-a-judge for infection prevention and control and antimicrobial resistance impact: comparing three main LLMs vs. human experts' assessment. Frontiers in Public Health 2026;14 View
  8. Pakull T, Bender N, Benson S, Fleischhauer A, Alsara M, Gromke T, Hilser T, Kaminski K, Pogorzelski M, Prasuhn N, Rosery V, Schadendorf D, Schuler M, Wiesweg M, Zaun G, Horn P, Friedrich C, Pretzell I. LLM-Generated Lay-Language Protocols for Molecular Tumor Board Patients: Evaluation of Quality and Clinical Usability. Journal of Medical Internet Research 2026;28:e99136 View
  9. Tan E, Yao W, Ong X, Foo J, Phua A, Abraham M. AI in Psychiatry for Improving Continuity of Patient Care: Protocol for a Mixed Methods Systematic Review. JMIR Research Protocols 2026;15:e95931 View
  10. Crawford A, Earle M, Castillo G, Nandlall N, Hilton M. Exploring Applications of AI in the Crisis Line Sector: Protocol for a Scoping Review. JMIR Research Protocols 2026;15:e95087 View
  11. Hao Z, Xu H, Wang C, Wang B, Qiu Z. Multimodal Digital Therapeutics Enhanced by Task Design and AI for Attention-Deficit/Hyperactivity Disorder Core Symptoms and Executive Functions in Children and Adolescents: Systematic Review and Network Meta-Analysis of Randomized Controlled Trials. Journal of Medical Internet Research 2026;28:e95043 View