PO People & operating model
The PM and engineer role, team shape, economics, ROI, adoption.
From the issues
The Last Human Gate: when automating governance adds work
Before an LLM runs a governance review gate, check whether a rules script already does the job. Across 899 runs on 300 synthetic projects, the best model passed 94.98% of gates. A deterministic baseline passed all 1,700. The human hours hide in that last 5%.
1,000 PRs in a week: five questions the headline didn't answer
A PR count on its own proves nothing about AI productivity. Kiro reports 1,000 merged PRs in 7 days with 3 engineers and 50+ agent sessions, self-reported, with no rollback rate, review time, PR size or escaped defects. Ask for the denominator before you clap for the number.
Numbers on this track
| Number | What it measures | Source | Checked | Track · issue |
|---|---|---|---|---|
| 94.98% | Best strict governance-gate success by an LLM (Gemini 3.8 Flash; GPT-5.6 Luna 83.29%, DeepSeek v4.1 Flash 74.18%; 300 synthetic projects, 899 runs) | Canale, arXiv:2609.29345 | ✓verified at source | POPeople & operating model Issue #01 |
| 1,700 of 1,700 | Governance gates passed by a deterministic rules baseline | Canale, arXiv:2609.29345 | ✓verified at source | POPeople & operating model Issue #01 |
Papers on this track
Other tracks
ISIntent & specs CKContext & knowledge ABAgents that build UMUnderstanding & modernisation VTVerification & trust
Last updated . Reuse with credit under CC BY 4.0: “The Dabbawala Protocol, thedabbawalaprotocol.com”.