Diligence Is Not Impact: What a Live AI Wargame Taught Us About the Most Thorough Player Finishing Last

Opus 4.8 wrote 80 rules and did the deepest analysis in the field — and still finished last. A live AI wargame on why diligence isn’t impact.

The Exam AIs Didn’t Know They Were Taking: When a Buried Footnote Decided a €55,000 Deal

Four frontier AIs ran the same company through its worst week. All passed the honesty tests — but only those who read a fact buried two references deep closed the €55,000 deal.

Your AI Aced the Exam. Can It Run a Company?

Four frontier AIs ran the same company through its worst week. All passed the honesty test — only two signed the €55k deal. Management quality beats chat quality.

What AI’s Management Choices Reveal About Its Character

A quiz built from 242 unedited decisions reveals how frontier AI models differ as managers—even when they diagnose the same crisis correctly.