Why are comparison queries the most stable query class? | Trakkr Research
Because they constrain the answer space more than open-ended best-of or general prompts. In the study, comparison queries reached 50.4% average agreement, the highest of the tracked query families.
Methodology: Built from 797,644 valid comparisons across 44,088 reports and 8 models, covering 6,439,133 model responses in the observed window.
Direct Answer
Mostly, because they constrain the answer space more than open-ended best-of or general prompts. Comparison queries reached 50.4% average agreement, the highest of the tracked query families.
What this means
Understanding query stability allows teams to allocate resources effectively, prioritizing content formats that either lock in consensus or exploit high-variance opportunities in the market.
Evidence table
| Metric | Value | Why it matters |
|---|---|---|
| Comparison-query agreement | 50.4% | Comparison prompts produce the highest average agreement. |
| General-query agreement | 42.2% | General prompts are less stable across models. |
Frequently Asked Questions
What is the average agreement rate for comparison queries?
Comparison queries reached a 50.4% average agreement rate across models.
How do general queries perform compared to comparison queries?
General prompts are less stable across models, showing a lower agreement rate of 42.2%.
What to do next
Related pages
Continue through the same study cluster.
- do ai models recommend the same brands - Related answer page
- how often is there perfect consensus across models - Related answer page
- high divergence prompts make up fourteen point six percent of the study - Related fact page
- cross model consensus tracker - Related tracker page
Data & Sources
- Same Question, Different AI, Different Answers - Flagship study behind this page
- Page JSON - Machine-readable companion file