Background <p>Accurate modality selection in pediatric imaging is critical, yet adherence to the American College of Radiology (ACR) Appropriateness Criteria remains limited. Large language models (LLMs) offer potential as decision support tools but often lack domain-specific accuracy.</p> Objective <p>To evaluate the performance of a locally run, context-aware chatbot (ped-Llama) based on an open-source LLM for providing personalized pediatric imaging recommendations grounded in ACR guidelines.</p> Materials and methods <p>A simulation-based study was conducted using 50 pediatric clinical scenarios derived from ACR guideline variants. The ped-Llama chatbot, built using a retrieval-augmented generation (RAG) approach with a locally deployed Llama-3.1–8B model, was compared against three generic LLMs (Llama-3.1–8B without RAG, Generative pre-trained transformer-4o [GPT-4o], Claude Opus) and three radiologists (junior resident, senior resident, specialist pediatric radiologist). Each provided imaging recommendations, which were classified as “Usually appropriate,” “May be appropriate,” or “Usually not appropriate” per ACR guidelines. Modal outputs were analyzed, and consistency across three independent LLM runs was assessed.</p> Results <p>ped-Llama achieved 80% (40/50) “Usually appropriate” recommendations, outperforming Llama-3.1–8B (46%), GPT-4o (54%), and Claude Opus (46%), and matching the specialist radiologist (76%). When both “Usually appropriate” and “May be appropriate” were accepted, ped-Llama achieved 90% accuracy. Consistency across runs was 72% for ped-Llama versus 44–50% for generic LLMs.</p> Conclusion <p>This study shows that a locally run, RAG-enabled chatbot based on an open-source LLM can provide guideline-concordant pediatric imaging recommendations with expert-level accuracy. Such systems offer a practical and scalable approach to artificial intelligence-assisted decision support in radiology.</p> Graphical Abstract <p></p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Locally deployed context-aware chatbot outperforms generic large language models for guideline-concordant pediatric imaging recommendations

  • Amit Gupta,
  • Krithika Rangarajan,
  • R G Krishna Kumar,
  • Somil Anshal

摘要

Background

Accurate modality selection in pediatric imaging is critical, yet adherence to the American College of Radiology (ACR) Appropriateness Criteria remains limited. Large language models (LLMs) offer potential as decision support tools but often lack domain-specific accuracy.

Objective

To evaluate the performance of a locally run, context-aware chatbot (ped-Llama) based on an open-source LLM for providing personalized pediatric imaging recommendations grounded in ACR guidelines.

Materials and methods

A simulation-based study was conducted using 50 pediatric clinical scenarios derived from ACR guideline variants. The ped-Llama chatbot, built using a retrieval-augmented generation (RAG) approach with a locally deployed Llama-3.1–8B model, was compared against three generic LLMs (Llama-3.1–8B without RAG, Generative pre-trained transformer-4o [GPT-4o], Claude Opus) and three radiologists (junior resident, senior resident, specialist pediatric radiologist). Each provided imaging recommendations, which were classified as “Usually appropriate,” “May be appropriate,” or “Usually not appropriate” per ACR guidelines. Modal outputs were analyzed, and consistency across three independent LLM runs was assessed.

Results

ped-Llama achieved 80% (40/50) “Usually appropriate” recommendations, outperforming Llama-3.1–8B (46%), GPT-4o (54%), and Claude Opus (46%), and matching the specialist radiologist (76%). When both “Usually appropriate” and “May be appropriate” were accepted, ped-Llama achieved 90% accuracy. Consistency across runs was 72% for ped-Llama versus 44–50% for generic LLMs.

Conclusion

This study shows that a locally run, RAG-enabled chatbot based on an open-source LLM can provide guideline-concordant pediatric imaging recommendations with expert-level accuracy. Such systems offer a practical and scalable approach to artificial intelligence-assisted decision support in radiology.

Graphical Abstract