My team and I were working on a different project and we faced a problem with being able to verify what an agent did, rather than relying on its own transcript.
It aims to capture items like shell commands, exit codes, file writes, tool calls, tests, and subagents. We build a timeline of the events but we do not store any prompts, responses, file contents, or tool outputs.
To get started, you can install it is as a Claude Code plugin or from the command line.
I would appreciate feedback on any edge cases as well as what execution items people want captured that we do not do already.