TensorX
返回文献探索

Paper · arXiv 2409.15127

Boosting Healthcare LLMs Through Retrieved Context

Jordi Bayarri-Planas, Ashwin Kumar Gururajan, Dario Garcia-Gasulla

19 upvotesSeptember 23, 2024arXiv 预印本
AI 摘要

The study enhances the factuality and reliability of large language models in healthcare through optimized context retrieval methods, achieving performance comparable to private solutions in benchmarks and proposing OpenMedPrompt for realistic open-ended answers.

Large Language Modelscontext retrievalhealthcare domainmultiple-choice question answeringopen-ended answersOpenMedPrompt

Abstract

Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language processing, and yet, their factual inaccuracies and hallucinations limits their application, particularly in critical domains like healthcare. Context retrieval methods, by introducing relevant information as input, have emerged as a crucial approach for enhancing LLM factuality and reliability. This study explores the boundaries of context retrieval methods within the healthcare domain, optimizing their components and benchmarking their performance against open and closed alternatives. Our findings reveal how open LLMs, when augmented with an optimized retrieval system, can achieve performance comparable to the biggest private solutions on established healthcare benchmarks (multiple-choice question answering). Recognizing the lack of realism of including the possible answers within the question (a setup only found in medical exams), and after assessing a strong LLM performance degradation in the absence of those options, we extend the context retrieval system in that direction. In particular, we propose OpenMedPrompt a pipeline that improves the generation of more reliable open-ended answers, moving this technology closer to practical application.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Boosting Healthcare LLMs Through Retrieved Context | TensorX