Skip to content
All threads

Thread

The Fabrication Problem

Most AI numbers are unverifiable. Source material fixes it. Self-checking catches only the surface. Trust signals are backwards.

Answered Source material collapses unsourced numbers to single digits. Prohibition outperforms monitoring 5x. Self-checking is unreliable because the same process generates and evaluates: it catches formatting and surface errors, not substance. Trust signals (citations, confidence, specificity) are higher in fabricated output than sourced output.

Open Does the fix work beyond reformulation tasks? Reasoning shows improvement (75% vs 38%), but strategy and creative untested.