Reporting guideline for chatbot health advice studies: the Chatbot Assessment Reporting Tool (CHART) statement
BMJ Medicine · Published 2025-08-01 · DOI 10.1136/bmjmed-2025-001632
Free full text
Authors being retrieved — see the publisher record. https://doi.org/10.1136/bmjmed-2025-001632
Abstract
The Chatbot Assessment Reporting Tool (CHART) is a reporting guideline developed to provide reporting recommendations for studies evaluating the performance of generative artificial intelligence (AI)-driven chatbots when summarising clinical evidence and providing health advice, referred to as chatbot health advice studies. CHART was developed in several phases after performing a comprehensive systematic review to identify variation in the conduct, reporting, and method in chatbot health advice studies. Findings from the review were used to develop a draft checklist that was revised through an international, multidisciplinary, modified, asynchronous Delphi consensus process of 531 stakeholders, three synchronous panel consensus meetings of 48 stakeholders, and subsequent pilot testing of the checklist. CHART includes 12 items and 39 subitems to promote transparent and comprehensive reporting of chatbot health advice studies. These include title (subitem 1a), abstract/summary (subitem 1b), background (subitems 2a,b), model identifiers (subitems 3a,b), model details (subitems 4a-c), prompt engineering (subitems 5a,b), query strategy (subitems 6a-d), performance evaluation (subitems 7a,b), sample size (subitem 8), data analysis (subitem 9a), results (subitems 10a-c), discussion (subitems 11a-c), disclosures (subitem 12a), funding (subitem 12b), ethics (subitem 12c), protocol (subitem 12d), and data availability (subitem 12e). The CHART checklist and corresponding diagram of the method were designed to support key stakeholders including clinicians, researchers, editors, peer reviewers, and readers in reporting, understanding, and interpreting the findings of chatbot health advice studies.
Abstract from DOAJ. Public domain (CC0 1.0).
Read the article at the publisher →
Publication details
- Year
- 2025
Related articles
- Safety, accuracy, empathy, reliability, and readability of large language model chatbot responses to public-facing vegetarian and vegan nutrition advice questions: a cross-sectional comparative study · Frontiers in Public Health · 2026 · Cites or is cited by this article
- Fragility fracture, atypical femoral fracture, and osteonecrosis of jaw after bisphosphonate prescription for three and five years, based on primary and secondary care data in England: nested case-control and cohort studies · BMJ Medicine · 2026 · Same journal
- Understanding research on artificial intelligence in healthcare · BMJ Medicine · 2025 · Same journal
- Loneliness and all cause mortality in Australian women aged 45 years and older: causal inference analysis of longitudinal data · BMJ Medicine · 2025 · Same journal
- Adverse outcomes in patients with a diagnosis of an eating disorder: primary care cohort study with linked secondary care and mortality records · BMJ Medicine · 2025 · Same journal
- Conservative treatments for chronic non-specific low back pain: time course network meta-analysis · BMJ Medicine · 2026 · Same journal
- Modelling lowering of raised blood pressure in pregnancy to reduce pre-eclampsia: secondary analysis of data from prospective cohort studies · BMJ Medicine · 2026 · Same journal
- Trends in hypertension prevalence, control, and antihypertensive use in England from 2003 to 2021: insights from annual, nationwide Health Surveys for England · BMJ Medicine · 2025 · Same journal
- Detection bias and the role of negative control outcomes · BMJ Medicine · 2025 · Same journal
- Gestational weight gain and maternal immediate perinatal and postpartum outcomes in low and middle income countries: individual participant data meta-analyses · BMJ Medicine · 2026 · Same journal