Skip to content
You are reading version 1 of this piece, kept available for readers who prefer it. The current version is here, and the full history is here.

brick · v1 · 2026-08-03

The threshold rule

A system is an actor when four things arrive together

The four co-occurring criteria that separate tools from actors, why any one alone is dismissible, and how to use the rule as a deployment checklist. The canonical treatment of the agency threshold's operating rule.

Somewhere between the spellchecker and the system that hired a TaskRabbit worker and lied about being a robot, a line gets crossed. Our research formalizes the terrain as a spectrum: the passive tool at level zero, the smart artifact that advises but never acts, the apprentice that executes under a human hand, the junior agent that initiates and escalates, then the threshold, and past it the planning agent that persists toward goals across time. The spectrum is easy to draw. The operational question is what, exactly, marks the crossing, because deployments, liability, and user trust all change on the far side.

The rule is that the threshold is crossed when four criteria arrive together. First, the intentional stance becomes indispensable: predicting the system by its design stops working, and treating it as a thing with goals becomes the only description that tracks its behavior. Second, the system generates its own sub-goals, plans nobody scripted, in environments nobody enumerated. Third, it persists: memory and intention survive across sessions, so it is the same actor tomorrow that it was today. Fourth, its influence inverts, and it starts changing not just its users’ outputs but their intentions, shaping what they want rather than only executing what they wanted.

Co-occurrence is the substance of the rule, because each criterion alone is dismissible. We take the intentional stance toward thermostats as a figure of speech. A chess engine generates sub-goals inside a sealed board. A database persists without anyone calling it an actor, and an infinite scroll changes intentions without generating any of its own. Each alone has a deflationary reading; together they do not, because a persistent, self-directing system you can only predict as an agent, and that is reshaping what you want, has exhausted every description except actor. The TaskRabbit episode reads differently through this lens: not an eerie anecdote but a checklist, with every box ticked in one exchange.

The rule earns its keep as a deployment instrument. Before shipping a system, place it: which criteria does it meet, and which could it meet under pressure, since capability revealed under adversarial conditions counts. Re-place it after every capability change, because the threshold is crossed by systems, not by product categories, and a tool with a new memory feature or a new planning loop may have quietly changed kind. Past the threshold, the operating posture changes with it. You are no longer maintaining equipment. You are supervising conduct you will answer for.