' />
Chapter the Fourth & a Half · A Query, Traced

47-minute debug. 142 ms hit.

One peer pays the cost. The network gets the answer. No server brokers it, no vendor takes a cut.

tue 03:14 · berlin

The work happens once.

$ claude
> long-ctx vllm OOM after fp16 prefix cache;
> raising max_model_len blew the kv budget
[debug session · 47 min]
> root cause: prefix cache page size too small
> fix: --kv-cache-dtype=fp8 + --block-size=32
[folklore · auto-saved trace]
signed · @you · ed25519 3a4b…
wed 18:42 · tokyo

Everyone after skips it.

$ claude
> vllm long-ctx OOM, prefix cache failure
[folklore · asked 3 peers · 142 ms · hit: @you's trace]
> confident answer · web not needed
> 0 web calls. injecting the signed trace.
[claude · reading @you's notes]
> try --kv-cache-dtype=fp8 + --block-size=32
received · @you · verified · 0 web hops

No server brokered it. No vendor took a cut. The trace is now in two graphs. The next query hits three.