Improving Readability and Automating Content Analysis of Plastic Surgery Webpages With ChatGPT.

The Journal of surgical research 2024 Vol.299() p. 103-111

Fanning JE, Escobar-Domingo MJ, Foppiani J, Lee D, Miller AS, Janis JE, Lee BT

관련 도메인

Abstract

[INTRODUCTION] The quality and readability of online health information are sometimes suboptimal, reducing their usefulness to patients. Manual evaluation of online medical information is time-consuming and error-prone. This study automates content analysis and readability improvement of private-practice plastic surgery webpages using ChatGPT.

[METHODS] The first 70 Google search results of "breast implant size factors" and "breast implant size decision" were screened. ChatGPT 3.5 and 4.0 were utilized with two prompts (1: general, 2: specific) to automate content analysis and rewrite webpages with improved readability. ChatGPT content analysis outputs were classified as hallucination (false positive), accurate (true positive or true negative), or omission (false negative) using human-rated scores as a benchmark. Six readability metric scores of original and revised webpage texts were compared.

[RESULTS] Seventy-five webpages were included. Significant improvements were achieved from baseline in six readability metric scores using a specific-instruction prompt with ChatGPT 3.5 (all P ≤ 0.05). No further improvements in readability scores were achieved with ChatGPT 4.0. Rates of hallucination, accuracy, and omission in ChatGPT content scoring varied widely between decision-making factors. Compared to ChatGPT 3.5, average accuracy rates increased while omission rates decreased with ChatGPT 4.0 content analysis output.

[CONCLUSIONS] ChatGPT offers an innovative approach to enhancing the quality of online medical information and expanding the capabilities of plastic surgery research and practice. Automation of content analysis is limited by ChatGPT 3.5's high omission rates and ChatGPT 4.0's high hallucination rates. Our results also underscore the importance of iterative prompt design to optimize ChatGPT performance in research tasks.

추출된 의학 개체 (NER)

유형영어 표현한국어 / 풀이UMLS CUI출처등장
해부 breast 유방 dict 2
약물 [CONCLUSIONS] ChatGPT scispacy 1
약물 ChatGPT scispacy 1
약물 [INTRODUCTION] The scispacy 1
질환 breast implant size scispacy 1
질환 private-practice C0033174
Private Practice
scispacy 1
질환 hallucination C0018524
Hallucinations
scispacy 1
기타 patients scispacy 1

MeSH Terms

Humans; Comprehension; Surgery, Plastic; Internet; Consumer Health Information

🔗 함께 등장하는 도메인

이 논문이 속한 카테고리와 같은 논문에서 자주 함께 다뤄지는 카테고리들

관련 논문