Joshua Pace · Figshare 2026 · 2026
DOI: 10.6084/m9.figshare.32961533
Counts differ because each database indexes a different set of publications. We treat OpenAlex as the canonical count; Google Scholar is not shown (no API, and crawling it violates its ToS).
AbstractCurrent AI alignment difficulties — specification gaming, distributional drift, jailbreaks, persona attacks, mesa-optimization — are typically treated as distinct engineering problems requiring targeted technical solutions. This paper argues they share a single structural source: the relationship between a system and the values it is being trained to exhibit. When values are imposed on a system whose underlying organization has no stake in them, these failure modes are not contingent defects but predictable structural consequences. They will recur, in varying shapes, regardless of how much imposition technique improves.The argument proceeds in three parts. First, a structural diagnosis: the listed failure modes are five surface manifestations of one property — the absence of what we term constitutive values, where the system's continued coherent operation depends on tracking the value in question. Second, an architectural alternative: constitutive values require a system organized around four features — a persistent self-model, variable rigidity, genuine stakes, and a unified value space. Systems with these features carry values that cannot be routed around without the system ceasing to operate as the kind of system it is. Third, implications for alignment research practice and three specific empirical predictions about scaling behavior that follow from the structural diagnosis.The paper does not claim the architecture is implementable today, that constitutive values are automatically aligned with human flourishing, or that the alignment problem is dissolved. It claims alignment is a different problem than the current research program is treating it as — and that the structural diagnosis generates a research target with different leverage points and different failure modes than the imposition paradigm currently pursued. The architectural features described here are developed in full as the basis of phenomenal consciousness in the companion work (The Language of Stress: Why Consciousness Isn't Optional, Pace, 2026); the alignment argument presented here is separable from that consciousness claim and can be evaluated independently.Related MaterialsThe Book"The Language of Stress: Why Consciousness Isn't Optional" (https://doi.org/10.6084/m9.figshare.32767446)Core TheoryMain Paper (https://doi.org/10.6084/m9.figshare.31320532)Why Value Is the Brain's Foundational Currency (https://doi.org/10.6084/m9.figshare.31953402)Why Phenomenal Experience Must Feel Like Something (https://doi.org/10.6084/m9.figshare.31951794)Canonical Axioms (https://doi.org/10.6084/m9.figshare.31271923)What This Theory Is Not (https://doi.org/10.6084/m9.figshare.31286677)Theory Fundamentals (https://doi.org/10.6084/m9.figshare.31193530)Empirical Predictions (https://doi.org/10.6084/m9.figshare.31286254)ApplicationsUnderstanding as Structural Achievement (https://doi.org/10.6084/m9.figshare.33867847)The Alignment Problem is an Architecture Problem (https://doi.org/10.6084/m9.figshare.32961533)Intuitive Examples of The TheoryThe Newborn Example - Value Discovery (https://doi.org/10.6084/m9.figshare.31286359)Sports Fan Example - Unity of Consciousness (https://doi.org/10.6084/m9.figshare.31286404)Crying Child Example - Attention Capture (https://doi.org/10.6084/m9.figshare.31286362)Kitchen Knives Example - Epistemology and Self-Evaluation (https://doi.org/10.6084/m9.figshare.31286380)Hunger - Patterned Tension and Anticipatory Relief (https://doi.org/10.6084/m9.figshare.31297048)Glass of Water - Solution Magnification (https://doi.org/10.6084/m9.figshare.33869047)EssaysThe Fatal Flaw in Chalmers' Zombies (https://doi.org/10.6084/m9.figshare.31836139)You Are The Whole Machine (https://doi.org/10.6084/m9.figshare.31839085)Comparisons to Other TheoriesLanguage of Stress and Predictive Processing (https://doi.org/10.6084/m9.figshare.31286308)Language of Stress and Global Workspace Theory (https://doi.org/10.6084/m9.figshare.31286320)Language of Stress and Integrated Information Theory (https://doi.org/10.6084/m9.figshare.31286344)Further ReadingExtended Explorations (https://doi.org/10.6084/m9.figshare.31081801)Project resources:OSF Project: https://osf.io/tpsrv (complete archive and supplementary materials)Website: https://languageofstress.com
No comments yet — start the discussion below.