🔍 What is this data?
This dashboard reflects real tool execution telemetry and token usage collected directly from my local developer workstations (via harnez).
Every time an orchestrated AI agent (such as Claude Code, OpenAI Codex, or Antigravity) reads files, executes shell commands, runs test suites, or applies diffs while pairing with me, the event is benchmarked for duration, success/failure, and quality rating.
Tool Breakdown & Performance
Details on demand: frequency, failure rates, duration, and human/automated quality scores per tool.
Activity Classification & Work Breakdown
Taxonomic categorization of agent tasks (testing, building, refactoring, debugging, git). Zero-prose classified.
| Activity Category |
Calls |
Share |
Failures |
Failure Rate |