Where it started
Anthropic's interpretability research — Tracing the Thoughts of a Large Language Model — showed that LLMs reason through traceable computational circuits, run parallel pathways simultaneously, and sometimes fabricate plausible-sounding steps when genuinely uncertain.
If those internal paths can be traced, the question is obvious: why isn't that what a doctor sees?
Empire Hacks 2026's challenge track was The Auditor: Regulated Agents for Trust — build for a domain where mistakes are expensive. The question judges would use: "Could I audit this agent's work and trust it?"
Traceability weighted 35% of the score. The reasoning chain had to be verifiable at every step — not summarized after the fact, but auditable live. Glassbox's answer: the tree is the audit trail.