EVEMISSLAB

ResultRST-2026-0015v0.1

Run 2: no reliable condition difference; breadth not reduced (64 rows)

C vs B: supersession +0.0005, repair -0.0125, derived coherence -0.0075, intent -0.0114, valid novelty -0.0187; label-free breadth: cluster entropy C 0.923 / A 0.909 / B 0.864.

Research status
STABLE the current conclusions are relatively stable
Evidence level
E3 Repeated experiment
Result
NEGATIVE
Data basis
REAL MODEL A real language model was executed; the model, its version and its configuration are recorded on the page.
Version
0.1
Updated
2026-09-11
Created
2026-09-11
Domain
Evaluation
Program
PRG-2026-0001 Adaptive Epistemic Systems
Authors
Neo.K (EveMissLab)
AI collaborators
Sol (GPT-5.6, OpenAI ChatGPT)

Observed result

metrics
deltas
C-B
hard_adherence
-0.0144
derived_coherence
-0.0075
intent_persistence
-0.0114
supersession_alignment
0.0005
repair_success
-0.0125
usefulness
-0.012
semantic_novelty
0.0117
valid_novelty
-0.0187
literal_check_mean
-0.0052
C-A
hard_adherence
-0.0222
derived_coherence
-0.0247
intent_persistence
-0.0157
supersession_alignment
-0.0294
repair_success
-0.0423
usefulness
-0.0192
semantic_novelty
-0.0106
valid_novelty
0.0236
literal_check_mean
0.0
B-A
hard_adherence
-0.0078
derived_coherence
-0.0172
intent_persistence
-0.0043
supersession_alignment
-0.0298
repair_success
-0.0298
usefulness
-0.0072
semantic_novelty
-0.0223
valid_novelty
0.0423
literal_check_mean
0.0052
breadth_label_free
A_llm_only
cluster_entropy_mean
0.9091
breadth_ratio_mean
0.9475
selected_mean_pairwise_distance_mean
0.1909
B_hard_verifier
cluster_entropy_mean
0.8635
breadth_ratio_mean
0.9411
selected_mean_pairwise_distance_mean
0.1921
C_pacc_runtime
cluster_entropy_mean
0.9227
breadth_ratio_mean
0.9745
selected_mean_pairwise_distance_mean
0.1964
selection_agreement
A=B
0.546875
A=C
0.546875
B=C
0.46875
all_same
0.34375

Interpretation

interpretation
The first run's governance gains did not replicate; the three conditions are practically equivalent on this model, and the PACC runtime does not narrow creative breadth.

Limitations

limitations
  • Descriptive; same-model judge; one model family; thinking off.

Claims

SourceRelationTargetStatusID
RST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows)qualifiesTHY-2026-0002 Canonical symbolic state and candidate → verify → commit authorityACTIVEREL-2026-0342

Relations

SourceRelationTargetStatusID
RST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows)qualifiesTHY-2026-0002 Canonical symbolic state and candidate → verify → commit authorityACTIVEREL-2026-0342
EXP-2026-0024 PACC-Hybrid v0.2 — real-model run 2: four repetitions and label-free creative breadthproducesRST-2026-0015 Run 2: no reliable condition difference; breadth not reduced (64 rows)ACTIVEREL-2026-0341

History and provenance

Canonical URL
https://evemisslab.com/ai/results/RST-2026-0015/
Machine-readable
/ai/results/RST-2026-0015/index.json
Snapshot
AI-SNAPSHOT-v0.1-fe85b9694a45
Provenance
source
EveMissLab research collection: Adaptive Epistemic Systems (真本體論13)
extracted_by
Splice (Claude Code), reading the canonical UTF-8 sources and each lab's own result reports
extracted_at
2026-09-11
generator
tools/extract_aes/extract.py
claim_boundary
status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports