Analysis & measurement
Instruments and their validity: model judges, platform panels, simulated agent societies. The longitudinal measurement campaign here is the weekly design's methodological precedent.
Entries are repositories — libraries, benchmarks, specifications — admitted as evidence, not as products.
Tasks in this category
- Calibrate model judges 0 of 91 entries
- Measure platforms longitudinally 13 of 91 entries
- Simulate agent societies & markets 0 of 91 entries
- Validate LLMs as instruments 0 of 91 entries
- Preregistration & power Planned 0 of 91 entries
- Psychometrics Planned 0 of 91 entries
- Missing data Planned 0 of 91 entries
Measure platforms longitudinally
Campaigns placed on this task
-
2026-08-03__platform-ecosystem-longitudinal-measurement
run type (the campaign's own designation) high-recall-map 13 of 91 entries ingested
- Ledger rows
- 42
- Screening rows
- 96
- Repos in ledger
- 16
Entries — 13 of 91 on this hub
-
Analysis & measurement Live
augurlabs/augur
Measure platforms longitudinally
Editorial draft Open-source community-metrics platform from the CHAOSS ecosystem; the platform-measurement campaign kept it as a historical schema comparator.
Forks 1005 Language Go Stars 692
Snapshot · retrieved UTC
-
Analysis & measurement Live
aveloxis/aveloxis
Measure platforms longitudinally
Editorial draft A community-metrics platform in Go; the platform-measurement ledger records it as the current form of the Augur lineage.
Forks 7 Language Go Stars 15
Snapshot · retrieved UTC
-
Analysis & measurement Unavailable
bohrdata/hfcommunity
Measure platforms longitudinally
Editorial draft A dataset project about a model-hub community. This namespace did not resolve at the 2026-08-10 snapshot (HTTP 404); the entry stays as a dated event, and the work's current home is listed separately.
Forks — Language — Stars —
Snapshot · retrieved UTC
-
Analysis & measurement Live
chaoss/community
Measure platforms longitudinally
Editorial draft The CHAOSS community handbook: metric definitions and governance for open-source health measurement. Admitted for its metric-definition receipts and ethics conventions.
Forks 190 Language JavaScript Stars 106
Snapshot · retrieved UTC
-
Analysis & measurement Live
chaoss/grimoirelab-perceval
Measure platforms longitudinally
Editorial draft Data-fetching component of GrimoireLab; pulls raw activity from many development platforms. Cited for its raw-versus-enriched data architecture.
Forks 187 Language Python Stars 324
Snapshot · retrieved UTC
-
Analysis & measurement Live
chaoss/grimoirelab-sortinghat
Measure platforms longitudinally
Editorial draft GrimoireLab's identity-management component. Cited alongside Perceval for architecture, with the campaign explicitly declining its identity-resolution features.
Forks 89 Language Python Stars 58
Snapshot · retrieved UTC
-
Analysis & measurement Live
ecosyste-ms/repos
Measure platforms longitudinally
Editorial draft An open dataset service indexing package registries and repositories; cited for schema and coverage diagnostics, with any data merge deferred pending terms and provenance.
Forks 15 Language Ruby Stars 72
Snapshot · retrieved UTC
-
Analysis & measurement Live
ghtorrent/ghtorrent.org
Measure platforms longitudinally
Editorial draft Site of the retired GHTorrent research dataset; kept as immutable historical evidence rather than a live backbone.
Forks 637 Language Ruby Stars 158
Snapshot · retrieved UTC
-
Analysis & measurement Unavailable
Nasif-Imtiaz-Ohi/BoDeGHa
Measure platforms longitudinally
Editorial draft Original namespace of a bot-detection classifier for GitHub accounts. This namespace did not resolve at the 2026-08-10 snapshot (HTTP 404); the entry stays as a dated event, and the maintained line is listed under its current address.
Forks — Language — Stars —
Snapshot · retrieved UTC
-
Analysis & measurement Live
package-url/purl-spec
Measure platforms longitudinally
Editorial draft The package-URL specification: canonical coordinates for software packages across ecosystems. Adopted by the campaign for dependency identity.
Forks 238 Language Python Stars 1090
Snapshot · retrieved UTC
-
Analysis & measurement Unavailable
RepoReapers/github-repo-dataset
Measure platforms longitudinally
Editorial draft Dataset namespace from the RepoReapers curation research. This namespace did not resolve at the 2026-08-10 snapshot (HTTP 404); the entry stays as a dated event, and its duplicate-cluster diagnostics survive in the scholarly record.
Forks — Language — Stars —
Snapshot · retrieved UTC
-
Analysis & measurement Live
sgl-umons/BoDeGHa
Measure platforms longitudinally
Editorial draft Maintained home of BoDeGHa, a classifier separating bot from human GitHub accounts; cited for its validation pattern.
Forks 13 Language Python Stars 27
Snapshot · retrieved UTC
-
Analysis & measurement Live
SOM-Research/HFCommunity
Measure platforms longitudinally
Editorial draft Current home of HFCommunity, a relational dataset built from a model hub's public activity; acquisition deferred to dated dumps.
Forks 2 Language Python Stars 16
Snapshot · retrieved UTC
Tasks with no entry today
6 of 7 tasks in this category carry no entry today. Each is listed with its state: ingested with no repository evidence in its ledger, on the record and not yet ingested, or planned and not yet run. Opening a row shows the campaigns placed on it, with their ledger and screening rows as recorded.
-
Calibrate model judges Not yet ingested 1 campaign on the record; one modern ledger has not been brought into the registry.
Campaigns placed on this task
-
2026-07-30__multi-model-contract-calibration
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 36
- Screening rows
- 56
- Repos in ledger
- 18
-
-
Simulate agent societies & markets Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-27__agent-markets-societies-simulations
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
SCOPE.md frozen before external discovery with comprehensive-continuation addendum; coverage claim: effort-bounded, v2 mechanism-class stop rule satisfied
-
2026-07-30__interactive-gamified-agent-worlds
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 28
- Screening rows
- 100
- Repos in ledger
- 31
-
-
Validate LLMs as instruments Not yet ingested 1 campaign on the record; one modern ledger has not been brought into the registry.
Campaigns placed on this task
-
2026-08-03__llm-as-measurement-instrument-validity-drift
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 48
- Screening rows
- 106
- Repos in ledger
- 4
-
-
Preregistration & power Planned No campaign has run yet. Grounded in crosswalk row C28 (internal crosswalk id).
-
Psychometrics Planned No campaign has run yet. Grounded in crosswalk row N03 (internal crosswalk id).
-
Missing data Planned No campaign has run yet. Grounded in crosswalk row N04 (internal crosswalk id).