SHARE measurement lineage · judge briefing
One question. Two instruments.
SHARE 2.0 keeps the transparent 25-check score, rebuilds Reusability from evidence available at deposit, and tests the same core question on the same reconciled records.
The original 1.0 archive had 183,872 records. Restricting it to the exact common cohort removed 29,303 negatives while retaining all 407 frozen legacy outcome labels.
The verdict first
Same records. Same SourceOf question. Cleaner separation between score and outcome.
Directness ledger
Every row says how comparable it really is.
“Direct” means the records, outcome labels, model family, and inference are held fixed. Refreshed and historical rows remain useful, but are labelled so different estimates are never presented as interchangeable.
Outcome continuity
The event lists moved, but they are recognizably the same construct.
Both versions use a public DataCite SourceOf relation indicating that a later object derives from the deposit. SHARE 2.0 refreshes the snapshot and normalizes relation direction.
The same 407 frozen legacy labels are present before and after the 183,872-to-154,569 cohort restriction. The 471 SourceOf, 114 verified-link, and 96 strict events below are distinct refreshed or narrower outcome definitions—not losses from the legacy 407.
Legacy inventory: 407
Refreshed SourceOf: 471
Verified link: 114
Strict three-year: 96
Association, honestly labelled
The positive signal persisted as the validation rules became more explicit.
odds per +10 deposit-time points
95% CI 4.97–6.61 · originally reported specification
odds per +10 SHARE 2.0 points
471 SourceOf records · 95% CI 1.76–6.90
odds per +10 SHARE 2.0 points
96 verified events · 95% CI 1.42–3.92
These three estimates answer closely related questions, but they are not a paired effect-size test. The completed matched analysis therefore places both instruments on the same 154,569 records and the same 407 labels: the archived 1.0 deposit-time score yields 5.097× odds per +10 (95% CI 3.161–8.219), while the 2.0 full score yields 4.167× (95% CI 2.078–8.355).
In the direct AUC analysis, SHARE 2.0's five-bucket model reaches 0.8786 versus 0.8209 for naive 25-signal counting. The paired improvement is +0.0577 (owner-clustered 95% CI +0.0020 to +0.1622), satisfying the prespecified primary decision. SHARE 1.0's common-cohort structured model remains higher at 0.9261; that less favorable version contrast is retained in the receipt rather than hidden.
Coverage expansion
Nine shared repositories, then two additions.
Direct comparison set · 9
Present in SHARE 1.0 and 2.0
New in live 2.0 · 2
Coverage added, not omitted from 1.0
This is why the live repository count is 11 while the submission-era static view consistently showed 9.
Evidence ledger
Follow every number to its governing source.
https://sharescore.org/methods-results
https://sharescore.org/framework/versions/2.0
https://sharescore.org/releases
https://api.sharescore.org/metrics/live
https://sharescore.org/metrics