Audience simulation

An audience simulator anticipates up to 90% of the criticism of a launch


Augur rehearses the reaction before a product or policy change goes out, and the same system goes from 0% to 73% accuracy depending on how it's asked.

September 26, 2026 · Translated from the Spanish original

What happened

Why it matters

The number

From 0% to 73% accuracy with the same weights, the same cases and the same evaluator.

Context

The inconsistency of answers depending on how the question is asked had already been measured for brands: two AI engines name the same brand in 81% and 43% of their answers. And the question of judgment, not capability, is the same one opened by Microsoft’s test where the best model gets 59.7% right.

What’s next

Bottom line

The paper measures a tool for rehearsing launches and ends up demonstrating something more useful: that much of what we call a model difference is a difference in wording.

Sources

Edited by Rodrigo Cornejo. How we select and verify: who writes these notes.

Related notes

← All notes