Skip to content

First-party measurement · September 12, 2026

How Often Two AI Answer Samples Disagreed

A reproducible analysis of 15 matched prompt-engine pairs showing answer, mention, competitor, and citation variation between two live samples.

Observed result

All 15 matched question-engine pairs produced different answer text. The mention verdict changed in 3 pairs, and the cited-domain set changed in 6 pairs.

Pair-level disagreement

Compared fieldPairs differingPairs compared
Answer text1515
VisiScan mention label315
Extracted competitor set215
Cited-domain set615

Method

Each pair holds the question and engine constant and compares sample 0 with sample 1. Exact answer strings, boolean mention labels, normalized competitor sets and normalized cited-domain sets are compared independently.

The scan used five questions and three available engines on one date. Its OTHER industry classification produced a weak prompt panel, so these rates describe this run only.