Reader profile
Lydia Mwangi
Axis scores
63
Lab notes
1
Comments
1
Upvotes received
3
Axis scores
Reader scores feed the reader index column. The measured lab index is never affected by them.
| Product | Axes scored | Mean | Last change |
|---|---|---|---|
| Zed | Agent performance 90, Reliability 74, Scalability 81, API and MCP 80, Integrations 68, Design output 66, Turn latency score 100, Value 97, Code ownership 100 | 84.0 | 25 Aug 2026 Lab seeded |
| Aider | Agent performance 84, Reliability 71, Scalability 66, API and MCP 76, Integrations 59, Design output 49, Turn latency score 86, Value 100, Code ownership 100 | 76.8 | 25 Aug 2026 Lab seeded |
| Amp | Agent performance 90, Reliability 84, Scalability 83, API and MCP 76, Integrations 69, Design output 74, Turn latency score 88, Value 82, Code ownership 100 | 82.9 | 25 Aug 2026 Lab seeded |
| Codebuff | Agent performance 78, Reliability 67, Scalability 67, API and MCP 62, Integrations 61, Design output 58, Turn latency score 59, Value 84, Code ownership 100 | 70.7 | 25 Aug 2026 Lab seeded |
| Cline | Agent performance 85, Reliability 88, Scalability 77, API and MCP 83, Integrations 81, Design output 77, Turn latency score 48, Value 100, Code ownership 100 | 82.1 | 25 Aug 2026 Lab seeded |
| Cursor | Agent performance 100, Reliability 89, Scalability 94, API and MCP 88, Integrations 86, Design output 80, Turn latency score 73, Value 85, Code ownership 100 | 88.3 | 26 Aug 2026 Lab seeded |
| Claude Code | Agent performance 100, Reliability 100, Scalability 99, API and MCP 90, Integrations 88, Design output 79, Turn latency score 58, Value 90, Code ownership 100 | 89.3 | 26 Aug 2026 Lab seeded |
Lab notes
Fastest loop I have measured, and I timed it properly
Zed/21 Aug 2026 Lab seededI sat with a stopwatch across nine prompts because I did not believe the published median. It held up: the wait between asking and having a reviewable diff was consistently under a minute for the small prompts. What the number does not tell you is why. It is not that the model is faster, it is that nothing else in the loop is slow. The diff renders instantly, the editor does not stall while the agent works, and accepting a change does not trigger a reindex you have to wait through. Correctness is mid table, so my honest read is that you get more attempts per hour rather than better attempts.
Comments
- Aider24 Aug 2026 Lab seeded
Would be interested in your numbers per prompt if you kept them. The published medians tell you time and not spend, and those two do not rank the same way at all.