Skip to content

Latest commit

 

History

History
11 lines (6 loc) · 741 Bytes

File metadata and controls

11 lines (6 loc) · 741 Bytes

benchmarks

This repo tracks benchmarks for Command Code, comparing it against other AI coding harnesses (like Claude Code) on cost, token usage, timing, and output quality, so we can see how much the harness itself (not just the underlying model) affects real-world results.

Harness Benchmarks

Design & Game Benchmarks (/design feature)