ExperimentEXP-2026-0104v0.1
Experiment B — binary vs numeric human measurement (declared)
Compare direct 0–10 rating, structured yes/no items and adaptive pairwise comparison on response time, missingness, inconsistency, test–retest, predictive validity and fatigue, with participants randomized across formats. Declared in the canonical index as the third v0.2 experiment; not designed in detail and not run.
Hypothesis
hypothesis- F3: well-designed binary/pairwise protocols beat direct numeric rating on at least one of response time, consistency, dropout, predictive validity or fatigue.
Procedure
procedure- Randomize participants across the three formats on the same artifacts; estimate latent quality with Bradley–Terry / IRT models; report measurement-process metrics alongside the estimates.
Runs
run_count- 0
Interpretation
interpretation- Declared, not run; listed so the program's falsifiable propositions each have their intended test on record.
Limitations
limitations- No protocol package exists yet; the canonical index gives the design in one paragraph.
Relations
| Source | Relation | Target | Status | ID |
|---|---|---|---|---|
EXP-2026-0104 Experiment B — binary vs numeric human measurement (declared) | tests | CLM-2026-0103 F3 — Binary burden hypothesis | ACTIVE | REL-2026-0492 |
EXP-2026-0104 Experiment B — binary vs numeric human measurement (declared) | tests | THY-2026-0107 Binary residual quality measurement (IBQF / BRQM) | ACTIVE | REL-2026-0493 |
History and provenance
- Canonical URL
- https://evemisslab.com/ai/experiments/EXP-2026-0104/
- Machine-readable
/ai/experiments/EXP-2026-0104/index.json- Snapshot
AI-SNAPSHOT-v0.1-fe85b9694a45- Provenance
source- EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by- Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at- 2026-09-11
generator- tools/extract_all.py
claim_boundary- status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports