Post

New research · Otolaryngology (ENT)
World journal of otorhinolaryngology - head and neck surgery · 1w
AI / informaticsWorld journal of otorhinolaryngology - head and neck surgery · 2026

Comprehensive Evaluation of AI Consent Forms in Otolaryngologic Surgery.

Sholem Hack, Rebecca Attal, Armin Farzad … Habib G Zalzal
Read paper
Otolaryngology (ENT)AI / informatics

Lay readers said AI-generated surgical consent forms adequately explained risks, benefits, and alternatives

Comprehensive Evaluation of AI Consent Forms in Otolaryngologic Surgery.

Sholem Hack … Habib G Zalzal
World journal of otorhinolaryngology - head and neck surgery · 2026
Background

Surgical consent documents are frequently written at reading levels exceeding average health literacy.

Purpose

This study evaluated the clarity, clinical accuracy, and acceptability of consent forms generated by GPT-4 and Claude for common otolaryngologic procedures.

Methods

A cross-sectional cohort of 300 English-speaking adults (15 raters per form) evaluated perceived clarity and signing comfort on 5-point Likert scales, and perceived trust using a binary (Yes/No) item, and completed eight binary quality assessments.

n = 10
Results
≥ 95%
share of readers who felt risks, benefits, and other treatment options were well explained
n = 10
More results

Mean lay ratings for clarity across AI-generated forms were high overall.

However, when analyzed at the form level, differences between models were not statistically significant for clarity (mean difference 0.04; t (9) = -0.80; p = 0.44) or signing comfort (mean difference 0.13, t (9) = -1.68, p = 0.13).

More results

Mean Flesch-Kincaid Grade Level was lower for AI-generated forms compared to official templates.

Although prompts targeted a 6th-8th grade reading level, achieved readability scores were slightly higher (8.8-9.4).

“
Conclusion · 1 of 2

In a non-clinical evaluation, AI-generated consent forms were perceived as clear and clinically complete, with model-specific trade-offs between perceived clarity and clinical detail.

Conclusion · 2 of 2

These perception-based findings-reflecting participant ratings of clarity, perceived trust, and willingness to sign rather than objective comprehension-are hypothesis-generating, and prospective clinical and legal validation in more representative patient populations is required.

Read paper
0 comments

No comments yet. Be the first.

Related papers

LatestFoundational
Randomized Trial
20.0
points lower score on a nasal symptom questionnaire with surgery than sprays
Randomized Trial
-1·60
dupilumab improved nasal polyp severity more than omalizumab
Animal / Preclinical
0%
no deep learning studies were tested with patients in real healthcare settings
Cross-sectional
21.1%
fear of illness was the most common workplace challenge for ear, nose, and throat staff
Guideline
“Lingual frenectomy may improve maternal pain during breastfeeding and may be an option in selected cases of phonetic alterations.”
AI / Informatics
4.50
expert doctors rated human-written reviews' overall scientific quality highest, above both AI chatbots
AI / Informatics
89.5%
Model correctly sorted healthy versus disordered voices in 89.5% of recordings
Cohort Study
51.2%
over half of children with a breathing tube experienced problems later on