Artificial intelligence as a modality to enhance the readability of neurosurgical literature for patients

J Neurosurg 142:1189–1195, 2025

The study evaluates ChatGPT 3.5 and GPT4’s ability to generate readable, accurate summaries of neurosurgical literature, enhancing patient comprehension. GPT4 showed higher readability and accuracy, suggesting its potential in improving patient education and bridging the gap between medical findings and public understanding.

Study Overview

Objective: Assess ChatGPT’s ability to generate readable, accurate neurosurgical summaries.

Methods: Analyzed 150 abstracts from top neurosurgical journals.

Models Used: GPT3.5 and GPT4.

Findings

Readability Improvement: GPT4 summaries more readable than original abstracts.

Scientific Accuracy: 84.2% of GPT4 summaries maintained moderate accuracy.

Readability Metrics: GPT4 outperformed GPT3.5 in multiple readability scores.

Implications

Patient Education: GPT4 can enhance neurosurgical literature comprehension for patients.

Health Literacy: Potential to improve health literacy nationwide.

Limitations and Future Research

Accessibility: GPT4’s restricted access limits broader application.

Future Studies: Explore GPT4’s use in other medical specialties.