理論THY-2026-0107v0.1
二元殘餘品質測量(IBQF/BRQM)
0–10 評分要回答者同時感知、建參照、校尺度、整合維度、映射成數字;被測狀態最重時負擔正好最高。BRQM 改成收集大量局部、具體、單一構念、非數值的二元或成對回答 b_i ∈ {0,1},讓測量系統用 Bradley–Terry/Thurstone/IRT 類模型重建潛在多維品質 θ̂_H,依「每單位人類成本的資訊增益」自適應選題,盲測與平衡設計,並明確保留評審分歧結構——因為二元觀測不是二元現象,分歧不是誤差。
定義
definitions- BRQM: 𝔔_H → {0,1}^N → θ̂_H; primitives b^abs ∈ {0,1} and b^pair ∈ {A, B}; skip = missing metadata, not a third value.
- Good-item conditions C_B = (local, single construct, concrete, temporally bounded, non-numeric).
- Adaptive selection i* = argmax E[IG_i] / C_H(i); stop when U_H < ε.
- Human residual object 𝔔_H^IBQF = (θ̂_H, Σ_H, N_obs, D_R, C_H, B_H, U_H, Grade_H); H-grades E uncontrolled rating … A+ cross-context validated.
前提
assumptions- A latent continuous quality exists behind local judgments (IBQF/FDCS micro-binary → macro-continuous emergence).
主張
claims- BinaryObservation ≠ BinaryPhenomenon; HumanObservation ≠ HumanScaleConstruction; NumericRating = State + ScaleUse + Context.
- MeasurementBurden ≠ MeasuredQuality; Disagreement ≠ Error; MeanPreference ≠ PreferenceStructure; Reliability ≠ Objectivity.
- HumanResidual ⇏ HumanOverridesFormalTruth — the hard gate is applied first.
形式化
formalization- P(A ≻ B) = σ(q_A − q_B) (Bradley–Terry); P(b_rij = 1) = σ(a_i θ_j − d_i + β_r), multidimensional λ_i^T θ_j, context-conditioned θ_j(c).
- C_rating = C_perceive + C_reference + C_scale + C_integrate + C_map; C_binary = C_local perceive + C_choose; C_binary, C_pair < C_rating is the hypothesis.
預測
predictions- With well-designed items, binary/pairwise adaptive protocols beat direct numeric rating on response time, consistency, dropout, predictive validity or fatigue in at least some settings (Falsifiable Claim 3).
證偽/失敗條件
falsification_conditions- If binary/pairwise protocols are worse than direct 0–10 rating on all of response time, consistency, dropout and predictive validity, the low-burden hypothesis must be revised.
已知限制
known_limitations- Not a clinical scale proposal; builds on EveMissLab's internal IBQF/MTF and FDCS theory (2025); Experiment B has not been run.
證據
| 來源 | 關係 | 目標 | 狀態 | ID |
|---|---|---|---|---|
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告) | tests | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0493 |
關係
| 來源 | 關係 | 目標 | 狀態 | ID |
|---|---|---|---|---|
THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | extends | THY-2026-0106 結構化品質、hard gate 與規格—驗證分離 | ACTIVE | REL-2026-0372 |
RES-2026-0102 品質線——不叫人打分數也能量成果品質 | develops | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0364 |
THY-2026-0108 高歧義成果的有型別、有版本品質本體 | extends | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0373 |
PAP-2026-0107 不要叫人類替自己的感覺打分數:IBQF 二元測量與低負擔品質評估 | formalizes | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0416 |
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告) | tests | THY-2026-0107 二元殘餘品質測量(IBQF/BRQM) | ACTIVE | REL-2026-0493 |
歷史與來源歷程
- Canonical URL
- https://evemisslab.com/ai/theory/THY-2026-0107/
- 快照
AI-SNAPSHOT-v0.1-fe85b9694a45- 來源歷程
source- EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by- Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at- 2026-09-11
generator- tools/extract_all.py
claim_boundary- status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports