Explicit probabilities can be trained and scored
Forecasting research supports calibration practice, teaming and systematic aggregation in appropriate domains. It does not prove every personal or business decision becomes better.
Evidence for individual ingredients is not evidence for the integrated protocol. We separate what is independently supported, what is a working hypothesis, what Experiment #001 measures, and what would force us to revise the system.
Forecasting research supports calibration practice, teaming and systematic aggregation in appropriate domains. It does not prove every personal or business decision becomes better.
Specific debiasing practices can change behavior beyond the exercise, but transfer is intervention-specific and should be measured rather than assumed.
Structured independent elicitation can preserve diverse signals; ordinary social influence can collapse the diversity that makes aggregation useful.
Lateral reading and provenance checks can improve evaluation of unfamiliar information sources.
Explicit update triggers are preferable to hoping we will remember to revise a plan later.
Systems and constraint models are useful when they generate testable implications rather than becoming explanatory dogma.
Repeated empirical support for the underlying mechanism or method.
Important ingredients are supported, while this exact transfer or use may be less directly established.
Useful for organizing complexity and generating interventions, but not treated as one complete predictive theory.
A lens that may reveal possibilities but should not outrank stronger evidence.
| Element | Protocol v1.2 |
|---|---|
| Participants | 6–8 selected participants |
| Foundation | Approximately 10–12 weeks; 6 in-person meetings |
| Foundation Gate | Reliability, decision-relevant candor and independent-first behavior must be working before monthly cruise |
| Cruise | One in-person Chartroom every 4 weeks + limited async panels |
| Async capacity | Approximately 3–4 Full Panels per group/month; typical member 2–4 responses/month |
| AI | Not used in the founding group decision process; future research layer only |
| Primary aim | Find observable decision/process signal worth testing further while keeping participant burden low |
What proportion of consequential cases can be resolved without a synchronous group meeting?
Does private-first async elicitation preserve meaningful differences before social influence?
Does the panel surface relevant observations and first-hand experience the owner did not already possess?
Do cases that reach the Chartroom actually contain more valuable disagreement, competing models or tacit information?
How much total participant time is required per resolved decision?
Does Foundation create enough trust for decision-relevant versions of sensitive real cases?
Can one in-person meeting every four weeks maintain the relationships/context needed for async cooperation?
Does the Scored Journal improve forecasting, review behavior and willingness to update?
Can the protocol function increasingly without an unusually strong founder-facilitator?
Once a stable human baseline exists, does any AI layer improve results enough to justify complexity and privacy cost?
A practical signal that async responsibility survives without constant facilitator effort.
Not “disclose everything”; disclose enough that omitted context is unlikely to reverse the judgment.
Genuine dissent when it exists; no requirement to manufacture disagreement.
Can people actually use the protocol on real decisions, and can we measure it?
Run more cohorts and separate robust patterns from founder/group effects.
Compare structured solo, async independent panels, ordinary discussion, structured room and later AI additions where appropriate.
Report useful results, null results and protocol changes so Reality Skill itself has an update history.
Mellers and colleagues; Good Judgment work; structured expert elicitation and IDEA/Delphi research.
Wisdom of crowds, hidden profiles, information pooling, conformity and collective intelligence.
Considering the opposite, premortem-style challenge, implementation intentions, calibration and outcome review.
Cognitive offloading, external representations, systems thinking, constraints, feedback, bottlenecks and value of information.
Keep the existing detailed source links below this section if you want to preserve the current bibliography verbatim.
In-person participation in Minsk is preferred when available because it offers richer human contact. Online participation exists for people who cannot reliably attend locally. We will still compare decision value, candor, trust and burden across delivery modes, but format remains a secondary implementation question rather than the identity of Reality Skill.