Run the Eval is an independent AI & tech review desk. We put the tools through real tasks, compare them head to head, and tell you what's actually worth paying for. The eval decides, not the marketing.

An antique brass balance scale with a blue Meta infinity loop on one pan and a purple Gemini star on the other, tipped toward Meta.

The Latest Evals

View all 62 →
Explainer

Meta Muse Is the AI Agent On-Ramp Entrepreneurs Needed

Meta's personal agent is free for most people, works while you sleep, and asks before it spends. Here's what X is saying, what it costs for real, what breaks, and where a one-person business should start.

Sep 14, 2026 · 7 min eval
How-To

How to Use Claude Code Hooks

Stop retyping the same instruction every session. Hooks turn 'please run the formatter' into a rule the tool can't forget — four steps, straight from the docs.

Aug 23, 2026 · 3 min eval
How We Test

The name is the method. We don't read the press release — we run the eval.

Read the full methodology →
01 Two sources or it doesn't ship Every load-bearing claim gets verified against at least two independent sources before publication.
02 Dead links are fabrications Every citation is machine-checked. A URL that doesn't resolve is treated as a made-up source and cut.
03 The adversarial pass A second review tries to refute every article before it ships. What can't be confirmed gets held.

Run the Eval is an independent AI & tech desk founded and edited by Micah Berkley, The AI Mogul — an AI evangelist and solutions architect who has spent over a decade building with and teaching practical AI. We don't read you the press release — we run the eval and show the receipts. More about the desk →

Get the verdict before the hype.

One email when a new eval ships. No spam, no sponsored picks — just receipts.

Unsubscribe anytime. We never share your email.