CAMS-CAN is a framework for scoring how well nations, companies, and cities are coordinating internally. It's published as open science because it's meant to be checked, not taken on faith. This page is a standing invitation to do exactly that: replicate a scoring pass, audit the rules, or find where it breaks.
Four ways in
Download the DIY kit, score an entity you know well, and compare it against a published dataset. Disagreement is useful data — tell us where and why.
The evidence gate, the change gate, the NA discipline, and the Node Value / Bond Strength formulas are all public. Find a logical gap or a case the rules don't actually cover.
An entity you expect the framework to score badly — a fast-moving crisis, an obscure polity, an edge case in the corporate or city mapping. Report exactly what happened, not just that it "felt wrong."
Run a multi-pass ensemble across model providers on a panel we haven't tested, and report ICC or an equivalent agreement statistic back. See the numbers we already have below, so you're extending them, not guessing at the bar.
What we've already found — check our work
Two independent panels have been run under the current scoring rubric, each as 18 formal scoring passes nested across three model families (Claude, Grok, Kimi). Full detail, caveats, and the formulas are on the DIY kit page; the framework's broader limitations are on Validation & Limits. The summary:
| Panel | Prompt | ICC(2,k) absolute | ICC(3,k) consistency |
|---|---|---|---|
| Australia, 2020–2025 | v1.1 | ||
| United States, 2020–2025 | v1.2-OPT |
What counts as a useful critique
We'd rather get one reproducible finding than ten impressions. A critique is easiest to act on when it's structural and specific.
Start here
Get the same four files used in every review above — project instructions and two blank schema templates, ready to paste into a Claude Project.
Get the DIY KitSend it in
Or open an issue or pull request directly against the source: github.com/KaliBond/wintermute. Everything here — framework, datasets, scoring tools, and this page — is licensed as open science under Common Property terms: fork it, run it, publish what you find.