Research

Energy balance

LLM responses had significantly higher scores than internet responses in the 'objectivity' and 'reproducibility' categories.

Practitioners may consider using LLMs for more objective and reproducible information on GLP1RA therapy.

StrongSupportsmedium confidence
LLM responses had significantly higher scores than internet responses in the 'objectivity' (mean 3.91, SD 0.63 vs mean 3.36, SD 0.80; mean difference 0.55, SD 1.00; 95% CI 0.03-1.06; P=.04) and 'reproducibility' (mean 3.85, SD 0.49 vs mean 3.00, SD 0.97; mean difference 0.85, SD 1.14; 95% CI 0.27-1.44; P=.007) categories.
Sarah Ying Tse Tan et al. · JMIR Formative Research · 2025

Why this rating

Based on the comparative study design involving evaluators and statistical analysis.

Source

Accuracy of Large Language Model Responses Versus Internet Searches for Common Questions About Glucagon-Like Peptide-1 Receptor Agonist Therapy: Exploratory Simulation Study

Sarah Ying Tse Tan et al. · JMIR Formative Research · 2025

DOI 10.2196/78289

otherCited 1×
Read the paper
DOI resolved against Crossref · corpus check 2026-06-10

This is one finding among thousands. Every one is graded and traced to its source, so you can see what the evidence actually supports. Browse the research →