Track
CK Context & knowledge
Context engineering, code knowledge graphs, retrieval over code, agent memory.
From the issues
Numbers on this track
| Number | What it measures | Source | Checked | Track · issue |
|---|---|---|---|---|
| >20% | Extra inference cost from repository context files (AGENTS.md), with no general gain in task success | Gloaguen et al., ETH Zurich, arXiv:2602.11988 | ✓verified at source | CKContext & knowledge Issue #01 |
| ~66% | Share of Cursor's system prompt trimmed | Cursor blog | self-reported | CKContext & knowledge Issue #01 |
| 7% | Cursor's total token reduction, "without degrading quality" (method not shown) | Cursor blog | self-reported | CKContext & knowledge Issue #01 |
| 22–25 of 27 | Top-four model scores on enterprise questions where the rules are stated | Era by Eon, arXiv:2609.30055 | ✓verified at source | CKContext & knowledge Issue #01 |
| ≤6 of 24 | Scores of four of six models when the answer hinges on a hidden fact in contradictory records | Era by Eon, arXiv:2609.30055 | ✓verified at source | CKContext & knowledge Issue #01 |
Papers on this track
- Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents? arXiv:2602.11988
- Era by Eon: Benchmarking Enterprise Agents on Hidden Knowledge arXiv:2609.30055
Other tracks
ISIntent & specs ABAgents that build UMUnderstanding & modernisation VTVerification & trust POPeople & operating model
Last updated . Reuse with credit under CC BY 4.0: “The Dabbawala Protocol, thedabbawalaprotocol.com”.