Research map: Model and Task-Aware Test-Time Scaling Strategies for Large Language and Vision-Language Models in Medicine: Evaluation Study

Back to the article

Papers in this map

  1. The Model Confidence Set · 2011 · 2146 citations · Cited by this paper
  2. Advanced deep learning methods for text generation in 2022-2024 literature: a systematic review · Artem V. Slobodianiuk · 2026 · Cites this paper
  3. How Social Media and Chatbot Bans Could Backfire · 2026 · Related
  4. Visual Instruction Tuning · Haotian Liu · 2023 · 1191 citations · Cited by this paper
  5. A Guideline-Concordant Chatbot Framework for Structured Colorectal Cancer Screening: Multistage Feasibility Study · 2026 · Related
  6. Training Language Models to Follow Instructions with Human Feedback · Long Ouyang · 2022 · 1182 citations · Cited by this paper
  7. Implementing Supported Digital Enhanced Cognitive Behavior Therapy for Binge Eating Disorder in Routine Care: Mixed Methods Service Evaluation · 2026 · Related
  8. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning · Daya Guo · 2025 · 1029 citations · Cited by this paper
  9. Development of a Blockchain-Based Platform to Enable Indigenous Data Sovereignty and Shared Research Participation With Indigenous Communities: Technology Prototyping and Community Engagement Study · 2026 · Related
  10. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing · Kristian Woodsend · 2025 · 779 citations · Cited by this paper
  11. China Has Moved to Regulate Expertise Online—and the West Should Pay Attention · 2026 · Related
  12. PubMedQA: A Dataset for Biomedical Research Question Answering · Qiao Jin · 2019 · 667 citations · Cited by this paper
  13. Behavior Change Content and Implementation of Large Language Model–Driven Conversational Agents in Cardiometabolic Care: Scoping Review · 2026 · Related
  14. What Disease Does This Patient Have? A Large-Scale Open Domain Question Answering Dataset from Medical Exams · Di Jin · 2021 · 597 citations · Cited by this paper
  15. A Supervised Fine-Tuned Large Language Model for Lifestyle Management in Patients With Prostate Cancer: Development and Evaluation Study · 2026 · Related
  16. Direct Preference Optimization: Your Language Model is Secretly a Reward Model · Rafael Rafailov · 2023 · 416 citations · Cited by this paper
  17. The Next Generation of Wearables Won’t Need the Cloud · 2026 · Related
  18. A generalist vision–language foundation model for diverse biomedical tasks · Kai Zhang · 2024 · 265 citations · Cited by this paper
  19. An Acceptance Criteria Framework for Determining the Implementation Fit of Custom Large Language Models in Public Health Interventions · 2026 · Related
  20. LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day · Chunyuan Li · 2023 · 229 citations · Cited by this paper
  21. Detecting Narcissistic Personality Disorder Traits on Forums: Proof-of-Concept Study · 2026 · Related
  22. Multi-Modal Understanding and Generation for Medical Images and Text via Vision-Language Pre-Training · Jong Hak Moon · 2022 · 213 citations · Cited by this paper
  23. OmniMedVQA: A New Large-Scale Comprehensive Evaluation Benchmark for Medical LVLM · Yutao Hu · 2024 · 71 citations · Cited by this paper
  24. Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale · Junying Chen · 2024 · 71 citations · Cited by this paper
  25. s1: Simple test-time scaling · Niklas Muennighoff · 2025 · 67 citations · Cited by this paper
  26. MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning · Jiazhen Pan · 2025 · 46 citations · Cited by this paper
  27. LlaVA-CoT: Let Vision Language Models Reason Step-By-Step · Guowei Xu · 2025 · 34 citations · Cited by this paper
  28. Benchmarking Large Language Models on Answering and Explaining Challenging Medical Questions · Hanjie Chen · 2025 · 26 citations · Cited by this paper
  29. Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models · Yuxiang Lai · 2026 · 23 citations · Cited by this paper
  30. Self-supervised multi-modal training from uncurated images and reports enables monitoring AI in radiology · Sang Joon Park · 2023 · 22 citations · Cited by this paper
  31. UltraMedical: Building Specialized Generalists in Biomedicine · Kaiyan Zhang · 2024 · 16 citations · Cited by this paper
  32. Towards Medical Complex Reasoning with LLMs through Medical Verifiable Problems · Junying Chen · 2025 · 15 citations · Cited by this paper
  33. MedCalc-Bench: Evaluating Large Language Models for Medical Calculations · Nikhil Khandekar · 2024 · 12 citations · Cited by this paper
  34. Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities? · Zhiyuan Zeng · 2025 · 5 citations · Cited by this paper
  35. Susceptibility of Large Language Models to User-Driven Factors in Medical Queries · Kyung Ho Lim · 2025 · 5 citations · Cited by this paper
  36. MedS³: Towards Medical Slow Thinking with Self-Evolved Soft Dual-sided Process Supervision · Shuyang Jiang · 2026 · 3 citations · Cited by this paper
  37. Untitled · Cited by this paper

Source: OpenAlex (CC0)