What this is: an illustration of what Latent feels like in the request path: token streaming, the internal-state probe scoring every token as it decodes, and the monitor deciding what to do. The token stream and per-token risk are the model's real generations and probe scores (from the evidence page); the latency figures and the pass / halt / steer / refuse / escalate actions are a product mockup of the intended behavior, not a benchmark.