How should real-world AI-tool proficiency be measured without turning usage into a fake expertise score?
I’m exploring a measurement problem rather than proposing that token count equals skill. I built a local-first technical alpha that records Claude Code and Codex activity, produces a signed privacy-sanitized snapshot, and separates activity telemetry f…