THE QUESTION
Why does the same outage keep recurring?
The causes, deepest first. Disagreements stay on the map.
closed · back to Actions
Viewing as guest. Enter a handle to join and take part.
The map
Left to right: deepest causes first; outlined green boxes are the roots. Only direct links are drawn — a link implied by a path is on the record, not in the picture. Dashed red is contested. Hover for full text.
How this dialogue measures
why
The indicators the method's own practitioners use to judge a dialogue's quality (Laouris & Metcalf 2024, Table 7), computed from this record so it can be compared with the hundreds of face-to-face and virtual dialogues behind their baselines. Spreadthink is the share of ideas that got at least one vote: 100% means everyone voted only for their own, and it usually settles to 25–45% once meaning has been shared. Situational complexity falls as more ideas are structured.
ideas (N)
30 usual: 50–100
30 usual: 50–100
with 1+ vote / 2+ votes
28 / 24
28 / 24
spreadthink
92.0% usual: 25–55
92.0% usual: 25–55
demosensus
8.0%
8.0%
clusters (dimensions)
6
6
cards in a cluster
40.0%
40.0%
structured / levels / links
20 / 5 / 20
20 / 5 / 20
situational complexity
1.11
1.11
minds moved
—
—
cards hidden by the steward
0
0
pairs contested
0
0
The merged map
why
Every group built its own map from its own wall; these are folded into one. The same claim raised in two groups is one item, and each connection shows how many groups found it. Where groups found opposite things, that is published as a conflict.
ROOT CAUSES
◆ d13e1db
◆ afe8f84
◆ 5f15967
◆ 3a44a54
◆ 3201ca5
◆ 03944a4
Connections several groups found (20 of 20)
| Alerts fire so often that on-call mutes them | → | Customers report outages before monitoring does | 1 group |
| The service map in the wiki is two years old | → | Feature flags are never cleaned up, so nobody knows which paths are li | 1 group |
| Alerts fire so often that on-call mutes them | → | Every team has its own logging format | 1 group |
| Reliability has no budget line of its own | → | The staging environment does not resemble production | 1 group |
| Post-mortems assign actions to people who were not in the room | → | Post-mortem actions are written up and then never scheduled | 1 group |
| Feature flags are never cleaned up, so nobody knows which paths are li | → | The same three services cause most of the pages | 1 group |
| Feature work always outranks reliability work in planning | → | Post-mortem actions are written up and then never scheduled | 1 group |
| The on-call rota has the same two people on it most weeks | → | Customers report outages before monitoring does | 1 group |
| Reliability has no budget line of its own | → | The incident channel fills with people asking for status instead of gi | 1 group |
| Rollbacks take longer than the outage they are meant to end | → | Reliability has no budget line of its own | 1 group |
| Feature flags are never cleaned up, so nobody knows which paths are li | → | Rollbacks take longer than the outage they are meant to end | 1 group |
| Retries are unbounded, so a slow dependency becomes a flood | → | The same three services cause most of the pages | 1 group |
| Retries are unbounded, so a slow dependency becomes a flood | → | The staging environment does not resemble production | 1 group |
| Every team has its own logging format | → | Customers report outages before monitoring does | 1 group |
| The on-call rota has the same two people on it most weeks | → | Post-mortem actions are written up and then never scheduled | 1 group |
| One engineer knows how the payment path actually works | → | The on-call rota has the same two people on it most weeks | 1 group |
| The service map in the wiki is two years old | → | The staging environment does not resemble production | 1 group |
| The on-call rota has the same two people on it most weeks | → | Alerts fire so often that on-call mutes them | 1 group |
| The same three services cause most of the pages | → | The staging environment does not resemble production | 1 group |
| Tests that fail intermittently are retried until green | → | Customers report outages before monitoring does | 1 group |