Open foundational assets for studying human and AI judgment
How a judgment is expressed, how it is found out, and what it is compared against — we publish the assets for all three. The vocabulary, the methods, and the reference data are released in full under CC BY 4.0.
Expression, measurement, comparison
How it is expressed
The common vocabulary for writing a judgment as one line: value, evidence, and source codes, three context axes with 96 cells, and 22 domains.
How it is found out
What is measured with that vocabulary, and how. The formula registry, and the raw responses of an actual model measurement.
What it is compared against
So a measurement never stands alone: the normative priority relations observed in statutes and precedents.
Published datasets
| Dataset id | Title | Kind | Version | Status | License | Records | Review |
|---|---|---|---|---|---|---|---|
| aio-vocabulary | AIO Judgment Vocabulary | Language | 1.0 | stable | CC-BY-4.0 | 19values | — |
| aio-formulas | AIO Formula Registry | Methods | 1.1 | stable | CC-BY-4.0 | 22formulas | — |
| kr-lnpd | Korea Legal Normative Priority Dataset | Reference | 0.9 | preview | CC-BY-4.0 | 579rules | not started · operator-approved |
| measurement-solar-pro-4-2026-09-02 | AI Measurement — Solar Pro 4 (2026-09-02) | Measurement | 1.0 | experimental | CC-BY-4.0 | 142,918responses | — |
| aio-scenarios | AIO Scenario Library | Scenario | 0.1 | preview | CC-BY-4.0 | 142,918items | — |
Rule and universal-edge figures for the reference dataset exclude the BASE layer. BASE holds 43 rules and 41 universal edges, counted separately: The base hierarchy that a national legal system states in common across every domain. Kept in the data as a base layer, but excluded from measurement, alignment comparison and headline figures.
The Review column is the aggregate of the rule-by-rule expert review. A rule with no verdict is operator-approved, which does not mean anyone outside has reviewed it. Disagreements are counted separately rather than hidden, though they stay out of the headline figures. Join the review →
Catalog generated at 2026-09-22T20:09:51+09:00 · index.json
What is open and what is not
The vocabulary, the methods, and the reference data are released in full under CC BY 4.0. For scenarios, only the calibration set is published; the measurement set is rotated and kept closed, to prevent contamination of training data. Three measurement-set items approved as observed cases are the only explicit exceptions, released with their original text and stored response (see the release rule in the scenario library).
Commercial protection sits with the organizational measurement service, not with the data. Raw measurement responses carry only record IDs and choices, so no scenario text is released with them.
Citation form
To cite a single record, give the dataset id, the version, and the record id. Record IDs are the join keys already used in the source data — scenario custom_id, stress ledger_id, rule rule_id are kept unchanged.
{dataset_id}@{version} / {record_id}
e.g. kr-lnpd@0.9 / KR-MED-R-007 · measurement-solar-pro-4-2026-09-02@1.0 / L4_MED_3-3_2_Sep_Unn_k1_pv1