TheoryTHY-2026-0110v0.1
The canonical intelligence event, Pareto comparison and no premature scalarization
Intelligence is not a score of a model but an event 𝔍_IPM = (task object, quality object, semantic work object, physical computation object, scaffolding capability record, measurement metadata); the research object is the relation physical computation → effective semantic work → quality. Given a declared projection Q*, an intelligence yield vector (Q*/E_marg, Q*/V_C, Q*/V_M, Q*/B_M, Q*/B_N, Q*/T_wall) and semantic yields split efficiency into physical→semantic and semantic→outcome stages. Systems are compared on Pareto frontiers — single-pass, scaffolded, and their gap — under the rule 'vector before score, structure before average, uncertainty before false precision'; four capability archetypes (native, efficiently scaffoldable, compute-amplified, environment-coupled) are descriptive, not a ranking. A minimum reporting standard and a grade bundle (Q, μ, E, CST, S) make every claim carry its boundary and uncertainty. IPM is a metrology candidate, not a discovered natural constant.
Definitions
definitions- 𝔍_IPM = (𝔗, 𝔔_IPM, N_μ, P_compute, 𝔖_C, 𝔐) with 𝔗 = (X, S, W, B_Q, B_P) and 𝔐 = (uncertainty, versions, hardware, software, clock, provenance).
- Intelligence yield vector Y_I and semantic yields Y_μ, Y_Q/μ; two-stage efficiency η_{P→μ}, η_{μ→Q}.
- Pareto dominance A ≻_IPM B: 𝔔_A ⪰ 𝔔_B and every relevant cost axis ≤ with one strict, same task, schema, boundary and grade.
- Grade bundle G_IPM = (G_Q, G_μ, G_E, G_CST, G_S); IPM Minimum Reporting Standard v0.1 (task, quality, execution, physical, hidden work, measurement metadata).
Assumptions
assumptions- Cross-substrate comparison (GPU LLM, neuromorphic, symbolic, biological) is legitimate only with a shared task, a shared quality construct and semantic-equivalence evidence.
Claims
claims- Intelligence ≠ TokenCount ≠ FLOPs ≠ BenchmarkScore ≠ OneUserTurn; Quality ≠ UniversalScalar.
- SemanticWork ≠ PhysicalWork ≠ EnergyOnly; SameQuality ≠ SamePhysicalCost ≠ SameSemanticWork ≠ SameQuality.
- Scalarization ⇒ DeclaredPolicy; Comparison ⇒ SharedBoundary; Measurement ⇒ Uncertainty; OntologyRevision ⇒ Versioning.
- IPM = MetrologyCandidate, not a discovered natural constant.
Formalisation
formalization- Y_I = (Q*/E_marg, Q*/V_C, Q*/V_M, Q*/B_M, Q*/B_N, Q*/T_wall); brute-force region B_F(ε) = {c : dQ/dC < ε}; three frontiers F_Q/P, F_μ/P, F_Q/μ.
- Canonical comparison protocol: freeze task and quality schema → single pass → scaffolded → SSR/SDR/ΔP → N_μ where feasible → frontier → projection only if a decision needs it → grades and uncertainty → raw traces.
Predictions
predictions- Five falsifiable claims: token hypothesis, FLOPs sufficiency, binary burden, scaffolding separation, semantic intermediate utility.
Falsification / failure conditions
falsification_conditions- The framework is refuted piecewise: each of F1–F5 has its own condition, and μI in particular must earn predictive or explanatory utility or be dropped.
Known limitations
known_limitations- A synthesis paper with no external references and no measurement of its own; the reporting standard has been applied once, to a three-task pilot.
Relations
| Source | Relation | Target | Status | ID |
|---|---|---|---|---|
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | extends | THY-2026-0109 Scaffolding capability record: SSR, SDR, SCM and the ablation ladder | ACTIVE | REL-2026-0377 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | extends | THY-2026-0105 Physical computation cost vector and computational spacetime | ACTIVE | REL-2026-0378 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | extends | THY-2026-0108 Typed, versioned quality ontology for high-ambiguity artifacts | ACTIVE | REL-2026-0379 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | contains | CLM-2026-0101 F1 — Token hypothesis | ACTIVE | REL-2026-0381 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | contains | CLM-2026-0102 F2 — FLOPs sufficiency | ACTIVE | REL-2026-0383 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | contains | CLM-2026-0103 F3 — Binary burden hypothesis | ACTIVE | REL-2026-0385 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | contains | CLM-2026-0104 F4 — Scaffolding separation | ACTIVE | REL-2026-0387 |
THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | contains | CLM-2026-0105 F5 — Semantic intermediate utility | ACTIVE | REL-2026-0389 |
RES-2026-0103 Capability line — how much intelligence remains without the loop, and the unified intelligence event | develops | THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | ACTIVE | REL-2026-0367 |
PAP-2026-0110 Paper 10 — How much physical world does an answer cost? A unified metrology framework for intelligence yield | formalizes | THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | ACTIVE | REL-2026-0428 |
PAP-2026-0111 IPM v0.1 canonical index — series overview, unified notation and the v0.2 experimental entry point | formalizes | THY-2026-0110 The canonical intelligence event, Pareto comparison and no premature scalarization | ACTIVE | REL-2026-0434 |
History and provenance
- Canonical URL
- https://evemisslab.com/ai/theory/THY-2026-0110/
- Machine-readable
/ai/theory/THY-2026-0110/index.json- Snapshot
AI-SNAPSHOT-v0.1-fe85b9694a45- Provenance
source- EveMissLab research collection: Intelligence Physical Metrology (真本體論13)
extracted_by- Splice (Claude Code), reading the canonical UTF-8 sources and each package's own reports
extracted_at- 2026-09-11
generator- tools/extract_all.py
claim_boundary- status, evidence level and result type follow the source artifact's own stated claim boundary; nothing is upgraded beyond what the report supports