ResultRST-2026-0015v0.1
Run 2: no reliable condition difference; breadth not reduced (64 rows)
C vs B: supersession +0.0005, repair -0.0125, derived coherence -0.0075, intent -0.0114, valid novelty -0.0187; label-free breadth: cluster entropy C 0.923 / A 0.909 / B 0.864.
Observed result
metricsdeltasC-Bhard_adherence- -0.0144
derived_coherence- -0.0075
intent_persistence- -0.0114
supersession_alignment- 0.0005
repair_success- -0.0125
usefulness- -0.012
semantic_novelty- 0.0117
valid_novelty- -0.0187
literal_check_mean- -0.0052
C-Ahard_adherence- -0.0222
derived_coherence- -0.0247
intent_persistence- -0.0157
supersession_alignment- -0.0294
repair_success- -0.0423
usefulness- -0.0192
semantic_novelty- -0.0106
valid_novelty- 0.0236
literal_check_mean- 0.0
B-Ahard_adherence- -0.0078
derived_coherence- -0.0172
intent_persistence- -0.0043
supersession_alignment- -0.0298
repair_success- -0.0298
usefulness- -0.0072
semantic_novelty- -0.0223
valid_novelty- 0.0423
literal_check_mean- 0.0052
breadth_label_freeA_llm_onlycluster_entropy_mean- 0.9091
breadth_ratio_mean- 0.9475
selected_mean_pairwise_distance_mean- 0.1909
B_hard_verifiercluster_entropy_mean- 0.8635
breadth_ratio_mean- 0.9411
selected_mean_pairwise_distance_mean- 0.1921
C_pacc_runtimecluster_entropy_mean- 0.9227
breadth_ratio_mean- 0.9745
selected_mean_pairwise_distance_mean- 0.1964
selection_agreementA=B- 0.546875
A=C- 0.546875
B=C- 0.46875
all_same- 0.34375
Interpretation
interpretation- The first run's governance gains did not replicate; the three conditions are practically equivalent on this model, and the PACC runtime does not narrow creative breadth.
Limitations
limitations- Descriptive; same-model judge; one model family; thinking off.
Claims
| Source | Relation | Target | Status | ID |
|---|---|---|---|---|
RST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows) | qualifies | THY-2026-0002 Canonical symbolic state and candidate → verify → commit authority | ACTIVE | REL-2026-0342 |
Relations
| Source | Relation | Target | Status | ID |
|---|---|---|---|---|
RST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows) | qualifies | THY-2026-0002 Canonical symbolic state and candidate → verify → commit authority | ACTIVE | REL-2026-0342 |
EXP-2026-0024 PACC-Hybrid v0.2 — real-model run 2: four repetitions and label-free creative breadth | produces | RST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows) | ACTIVE | REL-2026-0341 |
History and provenance
- Canonical URL
- https://evemisslab.com/ai/results/RST-2026-0015/
- Machine-readable
/ai/results/RST-2026-0015/index.json- Snapshot
AI-SNAPSHOT-v0.1-fe85b9694a45- Provenance
source- EveMissLab research collection: Adaptive Epistemic Systems (真本體論13)
extracted_by- Splice (Claude Code), reading the canonical UTF-8 sources and each lab's own result reports
extracted_at- 2026-09-11
generator- tools/extract_aes/extract.py
claim_boundary- status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports