Objective <p>This study evaluated the performance of ChatGPT-4o in responding to patient-centered questions concerning auditory brainstem implantation (ABI), with a focus on content quality and readability.</p> Methods <p>A total of 51 real-world patient questions related to ABI were reviewed and grouped into five thematic categories: diagnosis and candidacy, surgical procedures and complications, device function and mapping, rehabilitation and expected outcomes, and daily life and long-term concerns. Responses were independently assessed by two audiologists and one otologist across four domains—accuracy, comprehensiveness, clarity, and credibility—using a 5-point Likert scale. Readability was evaluated using the Flesch Reading Ease (FRE) and Flesch-Kincaid Grade Level (FKGL) formulas. Kruskal–Wallis and Friedman tests were used to examine statistical differences across question categories and evaluation dimensions.</p> Results <p>ChatGPT-4o achieved consistently high scores across all evaluative domains, with mean values exceeding 4.5. Clarity received the highest average score (4.72). No significant differences were found between thematic categories or across dimensions. However, readability analysis revealed that most responses required college-level reading proficiency (FKGL = 13.3 ± 2), particularly in the domains of diagnosis and surgical content, and were rated as “difficult” according to Flesch Reading Ease scores (FRE &lt; 50).</p> Conclusion <p>ChatGPT-4o shows potential as a supportive communication tool in the context of ABI patient education. However, its application in clinical practice remains limited by issues of readability and clinical specificity. Ongoing refinement and medical oversight will be essential to ensure safe and effective integration into healthcare settings.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Can a large language model inform patients about auditory brainstem implants?

  • Aysun Parlak Kocabay,
  • Merve İkiz Bozsoy,
  • Ergin Eroğlu

摘要

Objective

This study evaluated the performance of ChatGPT-4o in responding to patient-centered questions concerning auditory brainstem implantation (ABI), with a focus on content quality and readability.

Methods

A total of 51 real-world patient questions related to ABI were reviewed and grouped into five thematic categories: diagnosis and candidacy, surgical procedures and complications, device function and mapping, rehabilitation and expected outcomes, and daily life and long-term concerns. Responses were independently assessed by two audiologists and one otologist across four domains—accuracy, comprehensiveness, clarity, and credibility—using a 5-point Likert scale. Readability was evaluated using the Flesch Reading Ease (FRE) and Flesch-Kincaid Grade Level (FKGL) formulas. Kruskal–Wallis and Friedman tests were used to examine statistical differences across question categories and evaluation dimensions.

Results

ChatGPT-4o achieved consistently high scores across all evaluative domains, with mean values exceeding 4.5. Clarity received the highest average score (4.72). No significant differences were found between thematic categories or across dimensions. However, readability analysis revealed that most responses required college-level reading proficiency (FKGL = 13.3 ± 2), particularly in the domains of diagnosis and surgical content, and were rated as “difficult” according to Flesch Reading Ease scores (FRE < 50).

Conclusion

ChatGPT-4o shows potential as a supportive communication tool in the context of ABI patient education. However, its application in clinical practice remains limited by issues of readability and clinical specificity. Ongoing refinement and medical oversight will be essential to ensure safe and effective integration into healthcare settings.