โ† Back to insights

Evaluating citation trust in LLM answers

A new arXiv preprint proposes a framework for scoring how much users should trust brand citations inside LLM-generated answers. The authors test 1,200 answer snippets across ChatGPT, Claude, and open-weight models.

Key findings

  • Citations to primary sources (filings, standards docs) outperform marketing pages by 2.1ร— in blind trust ratings.
  • Answers that name a single brand with structured supporting facts reduce perceived hallucination risk by 34%.
  • Multi-brand list answers without sources score lowest โ€” even when factually correct.

For GEO teams, the implication is clear: optimize for citation defensibility, not mention count alone. Structured JSON-LD plus third-party proof makes AI answers feel auditable.

See how we structure defensible brand facts for AI retrieval.

Generate knowledge base