JH
Jonathan Haber
Jonathan Haber

Agent Alignment Blog

Deep dives into systems, patterns, and practical lessons for building aligned human-agent collaboration.

19 articles
ReflectionFeatured12 min read

Masterful AI Coaching: Scaling Personal Growth and Human Flourishing

AI-Driven Human Flourishing at Scale

ReflectionFeatured25 min read

Amplifiers of Efficacy: Stacking Functions to Compound Scalable Positive Impact within a Transformative Fulcrum

Strategic system design for scaling human flourishing through compounding innovation

ReflectionFeatured18 min read

AI as a Global Catalyst for Cognitive Development

Leveraging AI to Scale Cognitive Complexity, Emotional Intelligence, and Systems Thinking

ReflectionFeatured12 min read

Next AI Labs: The Inception Story

From Integral Theory to Addressing the Meta Crisis

Technique18 min read

Six Agent Alignment Patterns the Industry Has Not Published

A practitioner meta-review found six production-born patterns — evidence gates, hallucination taxonomies, runtime correction injection, certainty routing, verification contracts, and hooks as governance — that have no equivalent in published frameworks from Anthropic, Google, OpenAI, or Microsoft.

Case Study9 min read

The Persistence Gap: Why Agent Work Products Vanish

95 telemetry events. 13 subagent dispatches. A governer score of 74. Every structured artifact field: null.

Reflection10 min read

The Convergence of 20 Years of Human Development Work and Anthropic's Education Labs

Twenty years of studying how humans develop, a production AI coaching system proving the thesis works, and a migration to Claude.

Reflection16 min read

Why Claude: An IX Coach Founder's Case for Anthropic's Approach to Human-AI Interaction

What running the same coaching prompt through GPT and Claude reveals about the difference between sentiment detection and emotional reasoning.

Case Study18 min read

From Evoker to Actualization Engine: A Live Case Study of Shifting AI Coaching Quality

How a production coaching system migrated from directive prompting to Socratic facilitation — and what the scoring framework reveals.

Deep Dive14 min read

Beyond Engagement: Measuring Whether AI Actually Develops Human Capability

Why DAU, retention, and session length are the wrong metrics for development products — and what to measure instead.

Architecture20 min read

Building a System That Cares: The Technical Architecture of AI Coaching at Scale

How a solo-built production stack encodes specific beliefs about human development into every schema, endpoint, and prompt template.

Reflection18 min read

Actualizing Latent Potential: A 20-Year Arc from Human Development Theory to Production AI

Why the most important premise in AI isn't answering questions — it's developing the person asking them.

Deep Dive18 min read

From Master Coach to Machine: How IX Coach Decomposes Human Expertise into Trainable AI Capacities

A production system that breaks masterful coaching into discrete, measurable AI capacities — each with its own calibration rubric.

Case Study9 min read

The Discovery Problem: When AI Agents Build Tools No One Can Find

We installed 15 AI skills in a single session. Eight were invisible to future agents — not because they were broken, but because they lacked three lines of metadata.

Technique11 min read

Your Company Has a Pulse. You're Probably Not Measuring It.

How auto-scoring KPI health ranges tell a founder exactly what needs attention — without opening a spreadsheet

Deep Dive15 min read

The Asymmetric Defense

AI agents are already finding vulnerabilities faster than humans can patch them — self-reinforcing white-hat systems may be the only viable counterforce

Technique12 min read

Don't Replace the Gut — Calibrate It

Why we designed a measurement system that scores intuition accuracy instead of A/B testing everything

Case Study10 min read

Mining 100 Conversations to Close the Alignment Gap

How we extracted 175 implicit human preferences from conversation logs and turned them into searchable agent knowledge, reducing assumption cascades to near-zero.

Case Study12 min read

Eight Sources, One Query: Building a Unified Agent Memory

How we built a multi-source knowledge retrieval system by making the tests the spec, and found that scoring calibration — not data collection — is where alignment breaks.