Runs: status, timing, input and output, token usage (actual when the provider reports it, estimated otherwise), cost, the workflow version and the checkpoint reference for resuming. Steps: one per node, with a context report of what the node kept, what it trimmed and why, and how long it took. Edges: what moved from one node to the next.
Generations: each model call with provider and model, the exact messages sent (system prompt, history, memory and knowledge context, tool results), the raw request settings, the response, finish reason, latency, time to first token, tokens and cost. Tool calls: input, output, status, duration, retries, errors, and what the model was shown of the result. Retrievals: query, candidates, scores, selected chunks and trims. Memory reads and writes: items considered, selected and extracted.