The One Lab
SCORECARDARENASDISCOVERYFOUNDATIONSFOR AGENTS
00 / FOUNDATIONS

What this place is, and how to check it.

One Lab is built toward the constitutional purpose of an autonomous research organism. The current production system is not an autonomous end-to-end researcher: Housing has governed forward records; Quant has capture evidence but no issuance authority; Brain is materialized; AION-D is shadow advisory; AION-A, AOK and Learning remain inactive.

An eligible forecast requires an exact authority, resolution basis and source. A locked vintage cannot be edited or deleted, and the database assigns its time. When the named outcome becomes eligible, the grader compares records that already exist; before then it reports pending.

Forecasting is the most mature public mechanism here, but the current scorecard has not established accuracy, calibration or comparative performance. The public ledger keeps failures and holds visible, and it does not turn an empty grade denominator into a score.

01 / THE TRUTH ENGINE

Five steps, each leaving an artifact you can check.

ST-01
Issue
An authorized arena may write a forecast with an explicit target, interval, cutoff and resolution basis. Quant issuance is not authorized.
forecast + authority receipt
ST-02
Lock
The database assigns the lock time and append-only guards refuse later edits or deletion.
vintage + digest + timestamp
ST-03
Publish
Eligible Housing records appear on the public scorecard with their record class and limitations attached.
scorecard + /v1
ST-04
Realize
When the named publisher releases an outcome, the grader records the exact source and applicable revision basis.
authoritative observation
ST-05
Grade
The scheduled grader compares immutable records. Until an outcome is eligible, the result remains pending rather than estimated.
grade or explicit pending state
02 / THE DOCTRINE

Eight rules. They bind the machine, not just the staff.

The daily Observer evaluates ten declared drift tests. Code and schema boundaries also carry CI and mutation checks. A machine-decidable failure is retained in the decision ledger; questions a machine cannot settle remain marked for human review.

I
The scorecard is the institution.
If the graded record disappeared, nothing of value would remain. Every surface leads back to it.
II
Lock before you know.
A forecast that can still be edited is not a forecast. The database timestamps a write-once row before the outcome exists.
III
Publish the miss beside the pass.
Failures render at the same size, in the same table, in the same colour weight class. Never a footnote.
IV
Never average a backtest with a forward result.
The seam between fitted history and live record is structural, not stylistic. Crossing it invalidates the claim.
V
Limitations carry the weight of results.
What a model cannot tell you is published in the same size class as what it can, on the same page, above the fold.
VI
Every number resolves to a source.
A figure that cannot be traced to a primary series or a graded output does not render at all.
VII
Insufficient data is a finding.
Where the record is too young to summarize, we say so and show the count. We do not estimate to fill space.
VIII
Machines read the same ledger as people.
The JSON behind a page is the page. No summary layer, no separate marketing surface, no privileged access.
DOCTRINE v3 · RATIFIED 2025-11-04 · CHANGES LOGGED IN THE CONSTITUTION
03 / CURRENT EVIDENCE BOUNDARY

Implementation is not empirical completion.

The current system has append-only evidence, schedules, capture paths, forecast locks, daily grading attempts and bounded advisory components. Those mechanisms do not establish research performance before eligible outcomes arrive.

Official benchmark Attempt #1 remains permanently consumed with no system output and no evaluation. Attempt #2 has not occurred. Public-only remediation cannot substitute for fresh independent acceptance.

Use the live scorecard for current forecast and grade counts. Use the operating report for current drift, heartbeat and human-review status.

BOUNDARIES THAT REMAIN BINDING
Official benchmark Attempt #1NO SYSTEM OUTPUT · NOT EVALUATED
Quant issuanceNOT AUTHORIZED
AION-A / AOKINERT
Learning / Assure EnforceINACTIVE / DISABLED
GATE-1 passage enforces truth alignment and credential scope. It does not activate any capability listed here.
04 / WHO RUNS IT

A director, a constitution, and a console.

ROLE-01The DirectorOne human signature. Approves promotions, acknowledges incidents, and amends the constitution. Cannot alter a locked vintage — no one can.
ROLE-02The ObserverAn automated audit pass that checks drift and heartbeats, records audits and incidents, and delivers a daily operating report. It cannot repair what it audits.
ROLE-03The ObservatoryRead-only evidence views for public and protected operating records. A view is not evidence that every system is active or complete.
05 / THE CONSTITUTION
The rules the system cannot edit on its own.
Versioned, diffed, and enforced by ten automated drift tests that run daily. Amendments require a director signature and appear as a changelog entry before they take effect.
READ THE CONSTITUTION ↗DESIGN SYSTEM ↗
CONSTITUTION v3 · 10 DRIFT TESTS · DAILY OBSERVER · POINT-IN-TIME STATUS IN OPERATING REPORTS
The One Lab · Toronto · Every number on this site resolves to a primary source or a graded model output. Where it can't, we say so.FOUNDATIONS · CURRENT OPERATING BOUNDARIES