Model internals · Five-part route

From activation to agent action

An agent changed a file, but that does not mean the model directly touched disk. This series separates weights, activations, silent and visible reasoning, tool calls, runtimes, and environment results in time order, then asks which internal evidence supports causal claims.

Reading boundary: internal readouts are instruments, not word-for-word autobiography. Every article separates correlation, causal evidence from intervention, and interpretation of what a model may be doing.

Overview of model activations flowing toward an agent action

Reading route

Build the map, then read out, identify, and intervene

All five articles reuse one retention-policy task so every concept lands on an input, state change, tool call, or file result.