faf-cli v6 Claude Opus 4.8
1/9 → 8/9
The Context Bench
Every AI re-reads your repo and guesses. We measured it — the same questions answered cold, then with one file of context. Graded against the repo's own answer key.
.faf11 real runs · mechanically graded · sha-stamped. FAF don't lie.
npx faf-cli benchYour AI answers cold, then reads your .faf and answers again. You get the score — and a receipt.
.faf; grading is mechanical token-matching — no judge, no opinion.These runs are in-session — honest, self-reported. The authoritative leaderboard runs each model under controlled conditions.