Lux: Five episodes in, Hex. We've walked through the stone-throwing motif, the agenthood-versus-agency split, causation inside a layer, and the evidence posture. Time to take stock. What did the Throw preprint actually add to the emergence calculus toolkit? Hex: Like an inventory check. What's new on the workbench? Lux: Exactly. Five contributions, each doing a specific job. Let's walk through them. Hex: Contribution one? Lux: A definition. An agent is a theory object. Not a system with goals, not something with consciousness, not a biological organism by default. A maintained package — a closure that persists under budgets — whose interface variables make stable counterfactual differences at an induced scale. Hex: That's a definition you can actually test. If the viability kernel is empty — Lux: No agent. The package doesn't persist. If feasible empowerment is zero — Hex: No agency. The interface doesn't make a difference. Lux: The definition has failure modes built in. You know exactly when to say "this isn't an agent" — and you know why. That's rare for definitions of agency. Most of them require you to evaluate intentions or goals, which you can't measure from the outside. Hex: So the definition is structural, not psychological. You don't need to crack open the system's mind — you look at the viability kernel and the empowerment channel. Lux: Structural, operational, and testable. And notice what it doesn't require. No consciousness. No utility function. No biological substrate. A thermostat could count if it has a maintained package and a nontrivial action channel. A bacterium could count if the same conditions hold at its scale. The definition is substrate-agnostic. Tool number one on the workbench. Hex: Tool number two? Lux: The measurement suite. Three orthogonal metrics, each operationalizing a different face of agenthood. Packaging stability — you build an empirical endomap and measure its idempotence defect. If the defect is high, the macro description isn't behaving like a stable object. Viability — the greatest fixed-point kernel under support semantics. How many states can the agent maintain indefinitely while respecting budget constraints? And difference-making — feasible empowerment, channel capacity from budget-respecting actions to outside futures. Hex: And these three connect to the six primitives through the dictionary we covered last time. Lux: Every primitive gets a row. The dictionary is a lookup table from "which primitive is active" to "what signal should I see in the data." Turn off protocol holonomy and the horizon-dependent empowerment gap vanishes. Turn off maintenance and the packaging defect stays high. It's ablation-ready by design. Hex: Measurement with teeth. Not just "we computed this number" but "here's what happens when we break it." Lux: Which brings us to where you test it. Hex: Tool number three — the substrate itself? Lux: The test bench. A minimal finite ring-world where toggle switches realize all six primitives. Protocol holonomy through phase-dependent displacement. Accounting through ledger and repair. Constraints through feasibility gating. Identity through discrete sectors. Operator rewriting through a skill variable. Hex: And deliberately small. Lux: Small enough that every state is enumerable, every claim is auditable. The limitations section is explicit — and this matters — the ring-world is a minimal witness, not a claim about agency in the wild. The definitions are scale-agnostic; the exhibits are scale-specific. Hex: So the substrate proves that the definitions work. It doesn't claim to be the real world. Lux: Exactly the right reading. A proof of concept, not a model of biology. And the paper is explicit about this — the definitions are scale-agnostic, meaning they work at any size, but the ring-world is a minimal witness. It's the simplest possible stage where all six primitives can be toggled independently. That makes it clean for auditing — you can attribute every signal to a specific switch. Hex: Tool four — the guardrails. Lux: The evidence suite. Positive results first: repair collapses packaging defect from one to zero. Protocol holonomy creates horizon-dependent empowerment — the gap appears at horizon two and widens. Learning through the skill variable increases empowerment as the agent gets better. Each positive result maps to a specific primitive. Hex: But the nulls are what make it an evidence suite instead of a demo. Lux: Two null regimes. First: single-action system — only one move, empowerment is zero for all horizons. That's the trivial baseline. Second: exogenous schedule — toggles flip on a fixed pattern, empowerment looks high, but the agent didn't cause it. The null audit catches the fake. Hex: We keep coming back to the nulls. They're the honesty gauge. Lux: Without them, you don't know if your positive results are real or artifacts of mis-modeling. The paper treats baselines as first-class evidence, not as appendix material. The phrase is: "these baselines are part of the evidence suite, not afterthoughts." Hex: And tool five — the verification stamp. Lux: Hashed configs, an artifact auditor, one-command regeneration of the entire evidence suite, and a Lean lemma anchoring viability iteration as a greatest fixed point. Hex: The chain of custody for the evidence. You can trace any number back to the exact configuration that produced it. Lux: That's the core of it. Every experiment is traceable. Every figure maps to a config hash. The Lean lemma pins one critical claim to machine-checked logic. It doesn't formalize the whole framework — but it anchors the foundation. The viability kernel computation, which everything else depends on, has a verified mathematical basis. Hex: Five tools. Definition, measurement, substrate, guardrails, verification. That's a dense contribution list for one preprint. Lux: And they don't stand alone. Each tool connects back to the broader Six Birds theory program. Hex: How so? Lux: The Become paper showed route mismatch in large-eddy simulation — filtering a nonlinear PDE and then evolving gives a different answer than evolving and then filtering. That's the same holonomy diagnostic the Throw paper uses for protocol sensitivity, just in a continuous PDE domain instead of a finite ring-world. Hex: Same pattern, different substrate. The holonomy tool travels between ring-worlds and differential equations. Lux: And it works in both because the pattern is structural, not domain-specific. The Notch paper showed records as local notches — staged carriers that cost something to maintain, whose translation depends on protocol. That's the same stone metaphor. In Notch, the stone carries time. In Throw, the stone carries agency. Different application, same primitives. Hex: So the stone metaphor really does tie the program together. Notch it for time, throw it for agency. Lux: Same stone, two readings. And the core Six Birds paper proved the primitives themselves are forced by structure. Hex: The self-generation theorem again? Lux: Process soup, interface lens, bounded refinement — six closure mechanics fall out canonically. The Throw paper's dictionary isn't "we picked six features that seemed useful." It's "these six are the only ones that can arise from the structural requirements of maintaining a closure under limited access." The workbench slots are carved by the mathematics, not by preference. Hex: So the workbench isn't arbitrary. The tool slots are determined by the mathematics. Lux: And the paper is honest about what's left. Single-agent only — no social agency, no bargaining, no norms. Interface is assumed, not discovered from microdynamics. Empowerment is not a goal theory — high difference-making capacity doesn't mean good outcomes. The toy substrate is a witness, not a model of real organisms. Primitive coverage is uneven — protocol and maintenance are deeply exhibited, identity staging and operator rewriting are demonstrated but not fully ablated. Hex: That's a lot of limitations to volunteer. Lux: Which is itself a tool. By naming the gaps explicitly, the paper defines its own boundary. You know exactly where it stops claiming. Hex: And what comes next? Lux: Three directions flagged. Multi-agent settings where constraints and identity tokens encode commitments and norms. Moving from exact kernels to approximate learned models while keeping the audit posture. And exploring alternative causal proxies — risk-sensitive empowerment, reachability volumes, intervention-based effect sizes — to see when they agree or diverge. Hex: Tools added, gaps named, directions flagged. That's a clean handoff to whoever picks it up next. Lux: And that handoff is deliberate. The Throw paper extends the Six Birds workbench from objects and time into agency. Same primitives, new application. And the stone-throwing motif? It's still there. Make a reliable difference in the outside world — that's the test. Hex: Workbench extended. That's a wrap on this series.