Methodologies

How the work gets made

How the Hahn-Solo agentic pipeline actually runs — methodologies that produced the case studies.

Do the Skills Earn Their Keep?

Do the Skills Earn Their Keep?

The skills my agents rely on are knowledge artifacts — so I measure whether they earn their keep: coverage, reliability signals, and skill-to-outcome quality. Evidence over vibes.

MethodologyAgentic
The Clickdummy Is a Contract

The Clickdummy Is a Contract

I built UI from the prose spec and it drifted — stacked headers, missing decorative elements. The fix: treat the approved clickdummy as a binding contract, enforced by a screenshot diff before review.

MethodologyAgenticContent SDK
The Real-Tenant Probe

The Real-Tenant Probe

Move the moment of truth forward. Before deep planning, run a two-day probe against a real tenant — and let it cancel the PRD if the foundation isn't there. Cancellation is a successful outcome.

MethodologyAgenticMarketplace
The Pipeline Is a Partnership

The Pipeline Is a Partnership

I tried to run the agentic pipeline end-to-end on a four-SDK-surface PRD. Tests stayed green. The tenant didn't. The fix wasn't more autonomy — it was phasing the work, with checkpoints between.

MethodologyAgentic
Skills, For Real

Skills, For Real

I called them skills. The runtime never did. Three iterations, one cutover commit, and 46 markdown fragments finally became something a fresh agent could actually discover.

MethodologyAgentic
The Marketplace SDK Dogfood Loop

The Marketplace SDK Dogfood Loop

Ten numbered runs hardened Marketplace skills; Run 10 (PageShot) shipped first in Pages. Later apps dogfed the same learnings—patches, scale proof, or regression when skills already matched.

MethodologyMarketplaceAgentic