實驗EXP-2026-0104v0.1
Experiment B——二元 vs 數值的人類測量(已宣告)
比較直接 0–10 評分、結構化是/否題與自適應成對比較在反應時間、缺答、不一致、重測、預測效度與疲勞上的表現,受試者隨機分配到不同格式。canonical index 宣告為 v0.2 第三個實驗;尚未細部設計、尚未執行。
假設
hypothesis- F3: well-designed binary/pairwise protocols beat direct numeric rating on at least one of response time, consistency, dropout, predictive validity or fatigue.
程序
procedure- Randomize participants across the three formats on the same artifacts; estimate latent quality with Bradley–Terry / IRT models; report measurement-process metrics alongside the estimates.
執行
run_count- 0
詮釋
interpretation- Declared, not run; listed so the program's falsifiable propositions each have their intended test on record.
限制
limitations- No protocol package exists yet; the canonical index gives the design in one paragraph.
關係
| 來源 | 關係 | 目標 | 狀態 | ID |
|---|---|---|---|---|
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告) | tests | CLM-2026-0103 F3——二元負擔假說 | ACTIVE | REL-2026-0492 |
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告) | tests | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0493 |
歷史與來源歷程
- Canonical URL
- https://evemisslab.com/ai/experiments/EXP-2026-0104/
- 快照
AI-SNAPSHOT-v0.1-fe85b9694a45- 來源歷程
source- EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by- Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at- 2026-09-11
generator- tools/extract_all.py
claim_boundary- status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports