EVEMISSLAB
English

理論THY-2026-0107v0.1

二元殘餘品質測量(IBQF/BRQM)

0–10 評分要回答者同時感知、建參照、校尺度、整合維度、映射成數字;被測狀態最重時負擔正好最高。BRQM 改成收集大量局部、具體、單一構念、非數值的二元或成對回答 b_i ∈ {0,1},讓測量系統用 Bradley–Terry/Thurstone/IRT 類模型重建潛在多維品質 θ̂_H,依「每單位人類成本的資訊增益」自適應選題,盲測與平衡設計,並明確保留評審分歧結構——因為二元觀測不是二元現象,分歧不是誤差。

研究狀態
PRELIMINARY 已有初步的形式化或觀察
證據等級
E0 僅有概念
資料基礎
THEORY 純理論推理,沒有量測。
版本
0.1
更新
2026-09-02
建立
2026-09-02
領域
Cognitive Science
計畫
PRG-2026-0101 智能的物理計量(IPM)
作者
Neo.K (EveMissLab)
AI 協作
Aletheia (GPT-5.6 Sol, OpenAI ChatGPT)

定義

definitions
  • BRQM: 𝔔_H → {0,1}^N → θ̂_H; primitives b^abs ∈ {0,1} and b^pair ∈ {A, B}; skip = missing metadata, not a third value.
  • Good-item conditions C_B = (local, single construct, concrete, temporally bounded, non-numeric).
  • Adaptive selection i* = argmax E[IG_i] / C_H(i); stop when U_H < ε.
  • Human residual object 𝔔_H^IBQF = (θ̂_H, Σ_H, N_obs, D_R, C_H, B_H, U_H, Grade_H); H-grades E uncontrolled rating … A+ cross-context validated.

前提

assumptions
  • A latent continuous quality exists behind local judgments (IBQF/FDCS micro-binary → macro-continuous emergence).

主張

claims
  • BinaryObservation ≠ BinaryPhenomenon; HumanObservation ≠ HumanScaleConstruction; NumericRating = State + ScaleUse + Context.
  • MeasurementBurden ≠ MeasuredQuality; Disagreement ≠ Error; MeanPreference ≠ PreferenceStructure; Reliability ≠ Objectivity.
  • HumanResidual ⇏ HumanOverridesFormalTruth — the hard gate is applied first.

形式化

formalization
  • P(A ≻ B) = σ(q_A − q_B) (Bradley–Terry); P(b_rij = 1) = σ(a_i θ_j − d_i + β_r), multidimensional λ_i^T θ_j, context-conditioned θ_j(c).
  • C_rating = C_perceive + C_reference + C_scale + C_integrate + C_map; C_binary = C_local perceive + C_choose; C_binary, C_pair < C_rating is the hypothesis.

預測

predictions
  • With well-designed items, binary/pairwise adaptive protocols beat direct numeric rating on response time, consistency, dropout, predictive validity or fatigue in at least some settings (Falsifiable Claim 3).

證偽/失敗條件

falsification_conditions
  • If binary/pairwise protocols are worse than direct 0–10 rating on all of response time, consistency, dropout and predictive validity, the low-burden hypothesis must be revised.

已知限制

known_limitations
  • Not a clinical scale proposal; builds on EveMissLab's internal IBQF/MTF and FDCS theory (2025); Experiment B has not been run.

證據

來源關係目標狀態ID
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告)testsTHY-2026-0107 二元殘餘品質測量(IBQF/BRQM)ACTIVEREL-2026-0493

關係

來源關係目標狀態ID
THY-2026-0107 二元殘餘品質測量(IBQF/BRQM)extendsTHY-2026-0106 結構化品質、hard gate 與規格—驗證分離ACTIVEREL-2026-0372
RES-2026-0102 品質線——不叫人打分數也能量成果品質developsTHY-2026-0107 二元殘餘品質測量(IBQF/BRQM)ACTIVEREL-2026-0364
THY-2026-0108 高歧義成果的有型別、有版本品質本體extendsTHY-2026-0107 二元殘餘品質測量(IBQF/BRQM)ACTIVEREL-2026-0373
PAP-2026-0107 不要叫人類替自己的感覺打分數:IBQF 二元測量與低負擔品質評估formalizesTHY-2026-0107 二元殘餘品質測量(IBQF/BRQM)ACTIVEREL-2026-0416
EXP-2026-0104 Experiment B——二元 vs 數值的人類測量(已宣告)testsTHY-2026-0107 二元殘餘品質測量(IBQF/BRQM)ACTIVEREL-2026-0493

歷史與來源歷程

Canonical URL
https://evemisslab.com/ai/theory/THY-2026-0107/
機器可讀
/ai/theory/THY-2026-0107/index.json
快照
AI-SNAPSHOT-v0.1-fe85b9694a45
來源歷程
source
EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by
Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at
2026-09-11
generator
tools/extract_all.py
claim_boundary
status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports