Editorial notebook / 05 dispatches

Useful notes
for people in charge.

The notebook covers practical choices around AI workflows: scope, oversight, evidence, operations, and the hard parts that tool demos tend to skip.

Dispatch 01
Agent operations

8 min read

Stop asking agents for permission at every step. Start deciding where approval belongs.

Three approval models—per action, per plan, and per boundary—and a practical default for a small team running its first supervised agent pilot.

Read dispatch 01 Sources checked: 18 Feb–9 Apr 2026

Dispatch 02
Evaluation

9 min read

Before you add another agent, build a five-case evaluation suite.

A compact way to move beyond glossy demos: real tasks, observable outcomes, trace review, a failure case, and repeatability.

Read dispatch 02 Sources checked: 9 Jan 2026–22 Aug 2026

Dispatch 03
Workflow design

8 min read

The most useful first agent may be a workflow that does not look autonomous.

A decision guide for choosing the least-complex AI system that can improve a recurring task—and knowing when an agent is actually justified.

Read dispatch 03 Sources checked: 19 Dec 2024–22 Aug 2026

Dispatch 04
Operations

7 min read

Before you trust an agent, make its work easy to replay.

A practical review packet helps a team inspect the request, boundary, path, outcome, and human decision—without mistaking a polished summary for evidence.

Read dispatch 04 Sources checked: 9 Jan–23 Feb 2026; retrieved 22 Aug 2026

Dispatch 05
Context design

8 min read

Context is a budget. Spend it on the next decision.

A practical way to give an agent the smallest trustworthy packet for the job, then leave a handoff a human or fresh session can actually use.

Read dispatch 05 Sources checked: 29 Sep–26 Nov 2025; retrieved 22 Aug 2026

Dispatch 06
Independent test

10 min read

We spent $0.05 testing Jev. Its calibration held. The pipeline around it didn’t.

A pre-registered 1,000-call check of TypeSafe’s System One model: stated confidence tracked accuracy in every bin we could test, and it answered all 1,150 calls with typed output — while a chat baseline failed to return usable JSON on a third of calls.

Read dispatch 06 Self-measured + vendor claims checked: 15–22 Sep 2026

The format

One sharp question

Each note begins with a decision a team actually has to make.

The standard

Primary material first

External claims are sourced, limitations are stated, and the working opinion is distinguished from reported fact.

The promise

No content treadmill

A smaller collection of durable guides beats a stream of thin “AI news.”