A new arXiv preprint proposes a framework for scoring how much users should trust brand citations inside LLM-generated answers. The authors test 1,200 answer snippets across ChatGPT, Claude, and open-weight models.
Key findings
- Citations to primary sources (filings, standards docs) outperform marketing pages by 2.1ร in blind trust ratings.
- Answers that name a single brand with structured supporting facts reduce perceived hallucination risk by 34%.
- Multi-brand list answers without sources score lowest โ even when factually correct.
For GEO teams, the implication is clear: optimize for citation defensibility, not mention count alone. Structured JSON-LD plus third-party proof makes AI answers feel auditable.
See how we structure defensible brand facts for AI retrieval.
Generate knowledge base