EVEMISSLAB

ExperimentEXP-2026-0104v0.1

Experiment B — binary vs numeric human measurement (declared)

Compare direct 0–10 rating, structured yes/no items and adaptive pairwise comparison on response time, missingness, inconsistency, test–retest, predictive validity and fatigue, with participants randomized across formats. Declared in the canonical index as the third v0.2 experiment; not designed in detail and not run.

Research status
IDEA only a question or a concept so far
Evidence level
E0 Concept only
Result
INCONCLUSIVE
Data basis
NOT RUN Designed and validated as a harness; never executed with a real model.
Version
0.1
Updated
2026-09-02
Created
2026-09-02
Domain
Evaluation
Program
PRG-2026-0101 Intelligence Physical Metrology (IPM)
Authors
Neo.K (EveMissLab)
AI collaborators
Aletheia (GPT-5.6 Sol, OpenAI ChatGPT)

Hypothesis

hypothesis
F3: well-designed binary/pairwise protocols beat direct numeric rating on at least one of response time, consistency, dropout, predictive validity or fatigue.

Procedure

procedure
Randomize participants across the three formats on the same artifacts; estimate latent quality with Bradley–Terry / IRT models; report measurement-process metrics alongside the estimates.

Runs

run_count
0

Interpretation

interpretation
Declared, not run; listed so the program's falsifiable propositions each have their intended test on record.

Limitations

limitations
  • No protocol package exists yet; the canonical index gives the design in one paragraph.

Relations

SourceRelationTargetStatusID
EXP-2026-0104 Experiment B — binary vs numeric human measurement (declared)testsCLM-2026-0103 F3 — Binary burden hypothesisACTIVEREL-2026-0492
EXP-2026-0104 Experiment B — binary vs numeric human measurement (declared)testsTHY-2026-0107 Binary residual quality measurement (IBQF / BRQM)ACTIVEREL-2026-0493

History and provenance

Canonical URL
https://evemisslab.com/ai/experiments/EXP-2026-0104/
Machine-readable
/ai/experiments/EXP-2026-0104/index.json
Snapshot
AI-SNAPSHOT-v0.1-fe85b9694a45
Provenance
source
EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by
Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at
2026-09-11
generator
tools/extract_all.py
claim_boundary
status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports