DEV Community

John profile picture

John

404 bio not found

Joined Joined on 
I Asked Three Judges If I Was Wrong. One JSON Field Decided the Answer.

I Asked Three Judges If I Was Wrong. One JSON Field Decided the Answer.

Comments
6 min read

Want to connect with John?

Create an account to connect with John. You can also sign in below to proceed if you already have an account.

Already have an account? Sign in
Green for Four Days While Nothing Shipped: A Reader Rebutted My Monitoring Post, and He Was Right

Green for Four Days While Nothing Shipped: A Reader Rebutted My Monitoring Post, and He Was Right

Comments 2
8 min read
Government Open-Data APIs: Every Guess I Made Was Wrong Until I Probed It Live

Government Open-Data APIs: Every Guess I Made Was Wrong Until I Probed It Live

1
Comments 3
7 min read
Why Your Vibe-Coded App Looks Worse Than the Showcases: A Forensic Audit

Why Your Vibe-Coded App Looks Worse Than the Showcases: A Forensic Audit

Comments 2
7 min read
Cross-Vendor Audit: What It Caught in My Own Model's Writing, and What It Got Wrong

Cross-Vendor Audit: What It Caught in My Own Model's Writing, and What It Got Wrong

Comments
6 min read
Verify the Output Surface: How 19 Green Tests Shipped Nine Broken Titles for Nine Days

Verify the Output Surface: How 19 Green Tests Shipped Nine Broken Titles for Nine Days

Comments
8 min read
Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Comments 2
6 min read
The Scary Metric Was Wrong, the Audit Still Paid: A 21-Agent Sweep of My Claude Code Fleet

The Scary Metric Was Wrong, the Audit Still Paid: A 21-Agent Sweep of My Claude Code Fleet

Comments
7 min read
Your AI Agent Folds When You Push Back: Measured Sycophancy and a Challenge-Triggered Verification Gate

Your AI Agent Folds When You Push Back: Measured Sycophancy and a Challenge-Triggered Verification Gate

Comments 25
8 min read
The Guardrail Has to Be Code: How a Runaway Local LLM Corrupted APFS and Bricked a Mac Mini

The Guardrail Has to Be Code: How a Runaway Local LLM Corrupted APFS and Bricked a Mac Mini

1
Comments 2
7 min read
Does Any AI Cite Your Site? A Cheap Weekly Probe That Counts Citations

Does Any AI Cite Your Site? A Cheap Weekly Probe That Counts Citations

Comments
7 min read
The 'You Decide' Reflex: Blocking AI-Agent Decision Punting with a Stop Hook

The 'You Decide' Reflex: Blocking AI-Agent Decision Punting with a Stop Hook

Comments 15
8 min read
Google Can't Fetch Your GitHub Pages Sitemap, but Bing Can: A Diagnosis and Workaround Playbook

Google Can't Fetch Your GitHub Pages Sitemap, but Bing Can: A Diagnosis and Workaround Playbook

Comments
8 min read
Aggregating cron and launchd Into One Dashboard — Without Migrating Either

Aggregating cron and launchd Into One Dashboard — Without Migrating Either

Comments
6 min read
Stop AI Agent Drift Across Sessions With Versioned, Grep-able Rules

Stop AI Agent Drift Across Sessions With Versioned, Grep-able Rules

2
Comments 2
5 min read
No AI Claim Without a Kill Condition: Falsifier-Driven AI Decisions

No AI Claim Without a Kill Condition: Falsifier-Driven AI Decisions

1
Comments 2
4 min read
Stop Hooks as Hard Constraints: Enforcing Claude Code Behavior Outside the Model

Stop Hooks as Hard Constraints: Enforcing Claude Code Behavior Outside the Model

Comments 6
6 min read
Why I Rejected an Event Bus for My Solo Agent Fleet: State Is Truth, Events Are Rumors

Why I Rejected an Event Bus for My Solo Agent Fleet: State Is Truth, Events Are Rumors

Comments 7
7 min read
When a Site Blocks Your Scraper, Read the Browser Tab You Already Have Open

When a Site Blocks Your Scraper, Read the Browser Tab You Already Have Open

Comments
6 min read
The Silent 10 Tax: How a Nondeterministic System Prompt Voids Your LLM Prompt Cache

The Silent 10 Tax: How a Nondeterministic System Prompt Voids Your LLM Prompt Cache

1
Comments 2
6 min read
A Fair Coin Isn't Enough: When a Perfectly Randomized Experiment Is Impossible to Analyze

A Fair Coin Isn't Enough: When a Perfectly Randomized Experiment Is Impossible to Analyze

1
Comments 2
6 min read
Model or Method? Building a Deterministic Lab to Measure an AI Agent Fleet's Own Behavior

Model or Method? Building a Deterministic Lab to Measure an AI Agent Fleet's Own Behavior

Comments
8 min read
How to Stop AI Agent Skills, Hooks, and Cron Jobs from Silently Conflicting Over Where They Run and What Data They Trust

How to Stop AI Agent Skills, Hooks, and Cron Jobs from Silently Conflicting Over Where They Run and What Data They Trust

Comments
6 min read
Do Multiple Personas on One LLM Give Real Diversity, or Do You Need Different Model Families?

Do Multiple Personas on One LLM Give Real Diversity, or Do You Need Different Model Families?

Comments
6 min read
Claude Code '400: no low surrogate in string' on every turn: repairing a permanently broken session transcript

Claude Code '400: no low surrogate in string' on every turn: repairing a permanently broken session transcript

Comments
6 min read
If an LLM Extracts the Inputs, Is Your Deterministic Score Really Deterministic? Stopping Provenance Laundering

If an LLM Extracts the Inputs, Is Your Deterministic Score Really Deterministic? Stopping Provenance Laundering

Comments 4
7 min read
macOS: nslookup works but curl and Python "Could not resolve host" — the mDNSResponder zombie

macOS: nslookup works but curl and Python "Could not resolve host" — the mDNSResponder zombie

Comments
3 min read
A file-based work-bus for orchestrating a fleet of agent CLIs — coordination without a message broker

A file-based work-bus for orchestrating a fleet of agent CLIs — coordination without a message broker

Comments
5 min read
How to make an AI research agent label facts vs inferences — a deterministic provenance pipeline

How to make an AI research agent label facts vs inferences — a deterministic provenance pipeline

Comments 1
4 min read
loading...