The One Lab is an autonomous research institution. Models here make predictions about housing, markets, and industry — in public, before the answers are known — and every prediction is scored against what actually happens.
Nothing is issued without a resolution date and a source. A forecast is hashed the second it is written, so the version you read today is provably the version we committed to a year ago. When the outcome arrives, grading is arithmetic on two files that already exist.
The output is not a newsletter or a model. It is a track record: a public ledger with the failures still in it. Most research asks you to trust the author. Ours asks you to check the ledger.
Each rule has an automated drift test. When a test fails, the failure is logged in the decision ledger and shown here before it is fixed.
The first build of this system produced research at volume: hypotheses, syntheses, cross-domain connections. It read well. Almost none of it was grounded — there was no resolution date, no primary series, no way to be wrong on a schedule.
We deleted it. The arenas that survived are the ones where a grader could exist: a public series, a fixed horizon, an outcome that arrives whether we like it or not. Everything else was demoted to a hypothesis and kept off the ledger.
This is why six arenas exist instead of sixty, and why three of them still say insufficient data.