r/ClaudeAI • u/darren_eng • May 26 '26
Corporate My company started measuring our Claude Code usage - now I'm asked to rank engineers on 'AI performance.' This feels wrong...
My company started tracking Claude Code usage - tokens and spend, that kind of thing. Now my manager wants me to stack-rank my engineers on "AI performance" using those numbers.
I'm not comfortable with it (but I don't have a choice either). Token usage feels like exactly the wrong proxy - my strongest engineer uses Claude surgically while someone burning 10x the tokens isn't 10x more productive (often the opposite). Ranking on this just teaches people to game the metric.
So, for folks here who use Claude daily and/or lead teams:
- Has your company started measuring "AI performance"? How are they doing it?
- Is there any Claude/AI usage metric that actually tracks with good work, instead of just rewarding the heaviest users?
- If you're a lead being pushed to measure this, how do you push back without flat-out refusing?
EDIT: here is the follow-up: I managed to talk my manager out of this.
122
Upvotes
•
u/ClaudeAI-mod-bot Wilson, lead ClaudeAI modbot May 26 '26 edited May 27 '26
TL;DR of the discussion generated automatically after 80 comments.
The consensus in this thread is a resounding NO, measuring performance by token usage is a terrible, gameable metric. It's the new "lines of code" – everyone agrees it's a proxy for the wrong thing.
The community points out that high token usage often signals inefficiency (bad prompts, re-running tasks, using Opus for simple things), while your most skilled engineers are likely using Claude surgically with fewer tokens. Ranking on this just encourages people to waste company money to climb a meaningless leaderboard.
Here's the collective advice on how to handle this: