Endometriosis information generated by ChatGPT and Gemini: expert-rated accuracy and communicative adequacy
preprint
OA: green
CC0
AI-generated summary
Brazilian endometriosis experts evaluated ChatGPT and Gemini responses to patient and physician queries, assessing the technical accuracy and communicative adequacy of large language model-generated information about the condition.
One-sentence paraphrase of the abstract; not a substitute for reading it. No clinical advice. How this works
Abstract
Dataset from a cross-sectional comparative evaluation study of large language model responses about endometriosis, rated blindly by specialists from the Brazilian Endometriosis Society. Twenty questions — ten framed from the perspective of patients and ten from the perspective of gynecologists without specific experience in the disease — were submitted to ChatGPT (GPT-4o) and Gemini (1.5 Pro) in June 2024, in Brazilian Portuguese. The 40 responses were segmented into 190 thematic parts and evaluated for technical accuracy and communicative adequacy by 16 expert raters between October 2024 and February 2025. Approved by the Research Ethics Committee of the Federal University of São Paulo (protocol 6.657.855).
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2026) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- openalex
- last seen: 2026-09-21T06:00:58.944781+00:00
License: CC0
· commercial use OK