Independent Research
-
What Does It Actually Mean to Test an AI System?
A practical look at AI testing: baselines, repeatable tasks, external scoring, time and cost, limitations, and unsuccessful results.
-
The Problem with AI Agents That Agree with Each Other
Agreement among AI agents can hide shared assumptions. Explore why independent checks and external evidence matter more than consensus alone.
-
Can Multiple AI Agents Make Better Decisions Than One?
Can multiple AI agents improve decisions? Tim Hauptrief explores the research question, useful comparisons, and the evidence still needed.


