All evaluated windows
Current model only; PASS and VETO together.
−0.269N = 3648
95% CI [−0.300, −0.240]0.301N = 3648
LOWER IS BETTER0.237N = 3648
FROZEN, NEVER REFRESHED0.488N = 3648
0.490N = 3648
−0.002N = 3648
ATTESTATION OF FORECAST INTEGRITY · Nº 0002
SOMNIA SHANNON · TESTNET 50312
VERIFIABLE FORECAST RECORDER / V1 + V2
Bring any estimator. Proof·Edge freezes its probability and a contemporaneous market snapshot, seals both under a salted Keccak-256 commitment, and anchors the batch before on-chain expiry. There is still no minimum lead-time rule; the margin is measured, published and warned about instead of enforced. Pending and resolved records are published together, so losses remain beside wins.
0x243125583bf2dd8d7c8a33710bc8c1377e5c86f959ae6c288a0236a1203c55430x69b5e331…0beb2afb ↗0x253a60a7…1dd16417dreamdex-ec-oracle-follow-strike-adapterpublished/forecast-events.jsonl0 WINDOWS · 0 ROOTS · EXCLUDED FROM PROOF + SCORINGMIN 31 S · 31s · MEDIAN 285 S · 4m45s · 13 OF 3881 UNDER 90 S · LOW_LEAD IS A WARNING, NOT A VERDICT16 WINDOWS · PUBLISHED, NOT YET SCORED1 EVENTS · EXCLUDED FROM CANONICAL CHAIN · SEE LEDGER INTEGRITY REPORTMIT0x253a60a726a063c0e14acd10d7a206a0b82308a8bc703ced5304c79a1dd16417Primary reading · sealed version only
Current model only; PASS and VETO together.
−0.269N = 3648
95% CI [−0.300, −0.240]0.301N = 3648
LOWER IS BETTER0.237N = 3648
FROZEN, NEVER REFRESHED0.488N = 3648
0.490N = 3648
−0.002N = 3648
Current model only; first recorded gate ruling was PASS.
−0.020N = 667
95% CI [−0.042, +0.001]0.227N = 667
LOWER IS BETTER0.222N = 667
FROZEN, NEVER REFRESHED0.483N = 667
0.487N = 667
−0.004N = 667
Predicted p(YES) against the frequency the market actually resolved YES. N = 3648.
0x253a60a7…Every point is a probability that was anchored on Somnia before its market expired and scored against the outcome that arrived afterwards — the windows the current-model cards above score. No point here could be added, removed or reselected after the fact: the bins are whatever the sealed bytes already contained. Bars are 95% Wilson intervals on the observed frequency. All ten bins hold at least one forecast. (dashboard/app/forecast-data.json, key resolve_score.calibration).
Each row is derived from the model_hash already sealed inside its forecast payload.
| MODEL HASH | SAMPLE | N | MEAN p_AGENT | MEAN p_MARKET | BRIER A / M | SKILL | 95% BOOTSTRAP CI |
|---|---|---|---|---|---|---|---|
0x253a60a7…1dd16417CURRENT PRODUCTION | ALL EVALUATED | 3648USABLE | 0.488 | 0.490 | 0.301 / 0.237 | −0.269 | [−0.300, −0.240] |
| RISK-GATE PASS | 667USABLE | 0.483 | 0.487 | 0.227 / 0.222 | −0.020 | [−0.042, +0.001] | |
0x5b5b26d8…687b9ecbHISTORICAL VERSION | ALL EVALUATED | 6TOO SMALL | 0.414 | 0.362 | 0.174 / 0.147 | −0.189 | [−0.875, +0.120] |
| RISK-GATE PASS | 3TOO SMALL | 0.389 | 0.306 | 0.153 / 0.095 | −0.601 | [−0.885, −0.476] | |
0x6a7015d6…99257755HISTORICAL VERSION | ALL EVALUATED | 4TOO SMALL | 0.213 | 0.212 | 0.186 / 0.144 | −0.297 | [−1.560, +0.125] |
| RISK-GATE PASS | 3TOO SMALL | 0.210 | 0.175 | 0.046 / 0.039 | −0.186 | [−1.645, +0.189] | |
0xe43b1820…0e8e5639HISTORICAL VERSION | ALL EVALUATED | 9TOO SMALL | 0.374 | 0.280 | 0.206 / 0.180 | −0.144 | [−0.980, +0.261] |
| RISK-GATE PASS | 2TOO SMALL | 0.447 | 0.461 | 0.332 / 0.254 | −0.308 | [−0.341, −0.263] | |
0x914a3008…d73c3b34HISTORICAL VERSION | ALL EVALUATED | 8TOO SMALL | 0.416 | 0.510 | 0.344 / 0.232 | −0.481 | [−1.378, +0.198] |
| RISK-GATE PASS | 4TOO SMALL | 0.486 | 0.477 | 0.216 / 0.249 | +0.131 | [−0.177, +0.304] | |
0x61b7d39b…a7c8bb41HISTORICAL VERSION | ALL EVALUATED | 46TOO SMALL | 0.504 | 0.507 | 0.390 / 0.258 | −0.514 | [−0.769, −0.245] |
| RISK-GATE PASS | 10TOO SMALL | 0.576 | 0.547 | 0.299 / 0.273 | −0.095 | [−0.241, +0.073] | |
0xeb49a182…9e93d7c9HISTORICAL VERSION | ALL EVALUATED | 144USABLE | 0.467 | 0.497 | 0.299 / 0.235 | −0.270 | [−0.437, −0.103] |
| RISK-GATE PASS | 30TOO SMALL | 0.514 | 0.519 | 0.213 / 0.221 | +0.033 | [−0.073, +0.135] |
Context only. Never compare this total with a single model version.
−0.272N = 3865
MEAN p_AGENT 0.487 · p_MARKET 0.490 · BRIER A/M 0.301 / 0.237−0.020N = 719
MEAN p_AGENT 0.484 · p_MARKET 0.487 · BRIER A/M 0.227 / 0.222* Scores update from append-only resolution events. They use only the sealed agent probability and the market probability captured at commit; there is no backfill or historical repricing. The interval is a deterministic 1,000-resample, 95% bootstrap.
agent market · risk gating never filters the recorded sample
EXHIBIT § 2.1 — SELECTED WINDOW
0x0000…2132EVIDENCE DIGEST0x02fa337dc51e614ffcd79c2b975242ca09c6250e4f6121cce9a6dd922b737fa9
Do not trust this document.
Recompute it.
A clean clone reproduces commitments and Merkle proofs, checks every production root from the submitter, reads expiry and outcome from Shannon, and matches each receipt, block time, emitter, root, and leaf count. New-format anchors also bind the preceding ledger head. Late and pending records stay visible but unscored.
ROOT0x243125583bf2dd8d7c8a33710bc8c1377e5c86f959ae6c288a0236a1203c5543
ANCHOR TX0x69b5e331c87151395c7db05be090aac22e9d7a77e7885b825c001ac20beb2afb ↗
git clone --recurse-submodules https://github.com/Vastargazing/proof-edge.git
cd proof-edge
npm install --no-save
npm run check
npm run verify:log
npm run verify:chain
npm run verify:completeness
npm run verify:allThe five checks below are the ones npm run verify prints. This panel imports the same src/evidence-verifier.ts the CLI uses and reads Somnia Shannon directly over public JSON-RPC; only the transport differs.