Research map: Confident but unsupported: auditing large language models against supplied evidence boundaries in drug-induced liver injury assessment
Back to the article
Papers in this map
Survey of Hallucination in Natural Language Generation
· Ziwei Ji · 2022 · 4543 citations · Cited by this paper
National trends in hydrocephalus mortality in the United States, 1999–2024: a serial cross-sectional (ecological) trend analysis
· 2026 · Related
Large language models encode clinical knowledge
· Karan Singhal · 2023 · 3892 citations · Cited by this paper
Immune reconstitution and long-term survival outcomes in LTHIs and TPs living with long-term ART: a propensity score-matched cohort study in Xinjiang, China
· 2026 · Related
Large language models in medicine
· Arun James Thirunavukarasu · 2023 · 3892 citations · Cited by this paper
A multicenter cross-sectional study on the correlation between lower urinary tract symptoms and sleep status in female nurses aged 40 years and above
· 2026 · Related
Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models
· Tiffany H. Kung · 2023 · 3884 citations · Cited by this paper
Financial toxicity and distress among prostate cancer patients in Lebanon: a cross-sectional study
· 2026 · Related
Foundation models for generalist medical artificial intelligence
· Michael Moor · 2023 · 1947 citations · Cited by this paper
Clients’ views and experiences of individual placement and support: an employment program for people in treatment for substance use
· 2026 · Related
EASL Clinical Practice Guidelines: Drug-induced liver injury
· Raúl J. Andrade · 2019 · 1198 citations · Cited by this paper
Advances in cardiovascular risk assessment indicators in obstructive sleep apnea
· 2026 · Related
Drug-induced liver injury
· Raúl J. Andrade · 2019 · 839 citations · Cited by this paper
Large language models versus expert clinicians in optimizing Chinese patient education materials for temporomandibular disorders: a mixed-methods study integrating readability, accuracy, actionability, and cultural adaptability
· 2026 · Related
Evaluation and mitigation of the limitations of large language models in clinical decision-making
· Paul Hager · 2024 · 717 citations · Cited by this paper
Occupational noise exposure and cardiovascular disease risk factors in China: mediating role of hearing loss
· 2026 · Related
The TRIPOD-LLM reporting guideline for studies using large language models
· Jack Gallifant · 2025 · 530 citations · Cited by this paper
A framework to assess clinical safety and hallucination rates of LLMs for medical text summarisation
· Elham Asgari · 2025 · 354 citations · Cited by this paper
AASLD practice guidance on drug, herbal, and dietary supplement–induced liver injury
· Robert J. Fontana · 2022 · 293 citations · Cited by this paper
Diagnostic reasoning prompts reveal the potential for large language model interpretability in medicine
· Thomas Savage · 2024 · 265 citations · Cited by this paper
Medical large language models are vulnerable to data-poisoning attacks
· Daniel Alexander Alber · 2025 · 191 citations · Cited by this paper
Evaluating large language model workflows in clinical decision support for triage and referral and diagnosis
· Farieda Gaber · 2025 · 151 citations · Cited by this paper
Drug-induced liver injury severity and toxicity (DILIst): binary classification of 1279 drugs by human hepatotoxicity
· Shraddha H. Thakkar · 2019 · 143 citations · Cited by this paper
Improving medical reasoning through retrieval and self-reflection with retrieval-augmented large language models
· Minbyul Jeong · 2024 · 130 citations · Cited by this paper
Medical large language models are susceptible to targeted misinformation attacks
· Tianyu Han · 2024 · 59 citations · Cited by this paper
Probabilistic medical predictions of large language models
· Bowen Gu · 2024 · 42 citations · Cited by this paper
MEGA-RAG: a retrieval-augmented generation framework with multi-evidence guided answer refinement for mitigating hallucinations of LLMs in public health
· Shan Xu · 2025 · 41 citations · Cited by this paper
The long but necessary road to responsible use of large language models in healthcare research
· Jethro C.C. Kwong · 2024 · 30 citations · Cited by this paper
The need for guardrails with large language models in pharmacovigilance and other medical safety critical settings
· Joe B. Hakim · 2025 · 27 citations · Cited by this paper
BERT-based language model for accurate drug adverse event extraction from social media: implementation, evaluation, and contributions to pharmacovigilance practices
· Fan Dong · 2024 · 26 citations · Cited by this paper
Can large language models detect drug–drug interactions leading to adverse drug reactions?
· Justine Sicard · 2025 · 25 citations · Cited by this paper
Large Language Models for Adverse Drug Events: A Clinical Perspective
· Md Muntasir Zitu · 2025 · 20 citations · Cited by this paper
Assessing the performance of large language models in literature screening for pharmacovigilance: a comparative study
· Dan Li · 2024 · 16 citations · Cited by this paper
Evaluating large language models' performance in answering common questions on drug-induced liver injury
· Yinuo Dong · 2025 · 10 citations · Cited by this paper
Toward an Explainable Large Language Model for the Automatic Identification of the Drug-Induced Liver Injury Literature
· Chunwei Ma · 2024 · 8 citations · Cited by this paper
Drug induced liver injury: a clinical perspective
· Victor M. Navarro · 2009 · 5 citations · Cited by this paper
Toward Real-time Detection of Drug-induced Liver Injury Using Large Language Models: A Feasibility Study From Clinical Notes
· Thanathip Suenghataiphorn · 2025 · 5 citations · Cited by this paper
DILIrank 2.0: An updated and expanded database for drug-induced liver injury risk based on FDA labeling and a literature review
· AyoOluwa O. Olubamiwa · 2025 · 5 citations · Cited by this paper
Epistemic and ethical limits of large language models in evidence-based medicine: from knowledge to judgment
· Wenxiu Qi · 2026 · 3 citations · Cited by this paper
Source: OpenAlex (CC0)