Skip to content
Back to Blog
Blog series
8 posts · 76 min read

Reviewers & improvement

Reviewer agents, rollouts, operator corrections becoming versioned StrategyRules.

Share:XBSMRedditHNEmail
1
What Production Agent Runtimes Actually Teach ContextOS: Twelve Laws of a Governed Harness illustration
July 25, 2026·24 min read

What Production Agent Runtimes Actually Teach ContextOS: Twelve Laws of a Governed Harness

The strongest production agent runtimes converge on twelve architectural laws: compile context, persist state, separate authority from containment, make side effects resumable, treat approvals as typed interrupts, and promote learning only through evidence and replay.

2
Adaptive Agent Harnesses: Learn From Experience Without Letting Production Rewrite Itself illustration
July 24, 2026·10 min read

Adaptive Agent Harnesses: Learn From Experience Without Letting Production Rewrite Itself

MemoHarness shows that execution experience can improve the control layer around an LLM. Production systems need a stricter pattern: adaptive performance inside an immutable safety envelope.

3
Harness Improvement Loops Need Replayable Environments illustration
May 14, 2026·7 min read

Harness Improvement Loops Need Replayable Environments

Why harness improvement needs replayable episodes, bounded mutations, scorecards, source closure, and promotion gates.

4
Autotune the Harness: Baking the Improvement Loop into ContextOS illustration
May 12, 2026·11 min read

Autotune the Harness: Baking the Improvement Loop into ContextOS

How ContextOS treats autotune as a gated loop over traces, scorecards, replay sets, bounded candidates, approval, and rollout.

5
Building a Compliance Reviewer Agent in 60 Lines and a Golden Set illustration
March 15, 2026·6 min read

Building a Compliance Reviewer Agent in 60 Lines and a Golden Set

How to build a compliance reviewer agent with a typed verdict envelope, rubric, golden set, and change-control queue.

6
Building a Reliability Reviewer Agent: 70 Lines Past the Compliance One illustration
March 18, 2026·4 min read

Building a Reliability Reviewer Agent: 70 Lines Past the Compliance One

How to extend the reviewer pattern for reliability: timeouts, retries, idempotency, fallback behavior, and rollback declarations.

7
Pack Rollout in Five Stages: Shipping a Context Pack Without Blowing Up Production illustration
April 11, 2026·7 min read

Pack Rollout in Five Stages: Shipping a Context Pack Without Blowing Up Production

A five-stage rollout model for Context Packs: shadow, internal, low-risk, monitored expansion, full release, and rollback.

8
From Operator Correction to Released StrategyRule: The Improvement Loop, Coded illustration
April 15, 2026·7 min read

From Operator Correction to Released StrategyRule: The Improvement Loop, Coded

How one operator correction becomes a reviewed, replayed, versioned StrategyRule that prevents repeat agent failures.