vix.ing · top · new · best · stats

AR-Med: Automated Relevance Enhancement in Medical Search via LLM-Driven Information Augmentation

2025/12/03 by Chuyue Wang, Wang, Chuyue, Jie Feng +9
Computer Science · Medicine · #Artificial Intelligence in Healthcare and Education #Benchmark (surveying) #Blueprint #Bridge (graph theory) #Computation and Language (cs.CL) #Domain (mathematical analysis) #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning in Healthcare #Multimodal Machine Learning Applications #Relevance (law) #Scalability #Scheme (mathematics) #Service (business) #Subject-matter expert

paper · pdf · doi:10.48550/arxiv.2512.03737

published in arXiv (Cornell University) (Cornell University)

openalex publication_date 2025/12/03 · openalex created_date 2025/12/05 · openalex updated_date 2026/07/28

Abstract

Accurate and reliable search on online healthcare platforms is critical for user safety and service efficacy. Traditional methods, however, often fail to comprehend complex and nuanced user queries, limiting their effectiveness. Large language models (LLMs) present a promising solution, offering powerful semantic understanding to bridge this gap. Despite their potential, deploying LLMs in this high-stakes domain is fraught with challenges, including factual hallucinations, specialized knowledge gaps, and high operational costs. To overcome these barriers, we introduce AR-Med, a novel framework for Automated Relevance assessment for Medical search that has been successfully deployed at scale on the Online Medical Delivery Platforms. AR-Med grounds LLM reasoning in verified medical knowledge through a retrieval-augmented approach, ensuring high accuracy and reliability. To enable efficient online service, we design a practical knowledge distillation scheme that compresses large teacher models into compact yet powerful student models. We also introduce LocalQSMed, a multi-expert annotated benchmark developed to guide model iteration and ensure strong alignment between offline and online performance. Extensive experiments show AR-Med achieves an offline accuracy of over 93%, a 24% absolute improvement over the original online system, and delivers significant gains in online relevance and user satisfaction. Our work presents a practical and scalable blueprint for developing trustworthy, LLM-powered systems in real-world healthcare applications.

Citations

Related