In the style of “this meeting could have been an email”

AI models don't recognize their own code and prefer the longer text


Fifteen model-benchmark combinations landed between 49% and 58% accuracy, and the choice is explained by the length of the answer, not by its author.

September 26, 2026 · Translated from the Spanish original

What happened

Why it matters

The number

r = 0.93 between the evaluator’s accuracy and how often its own solution was the longer one.

Context

The debate over what happens when content looks machine-made usually assumes the mark of origin is detectable. The alternative being built goes another way: a credentials checker that confirms the signature without proving the content.

What’s next

Bottom line

Anyone selling AI-generated content detection based on a model recognizing itself has a methodological problem before a product one: the signal it measures is the length of the text.

Sources

Edited by Rodrigo Cornejo. How we select and verify: who writes these notes.

Related notes

← All notes