Right. Spans show what got called, not what the agent thought it was doing in between. Ollie walks the causal chain to find where it broke, but intent still gets pieced together from outputs. Logging the model's reasoning as its own span helps, but most frameworks skip it by
Spans show outputs not intent; logging reasoning helps
By
–