Oliver Christian Neutert · Zenodo (CERN European Organization for Nuclear Research) 2026 · 2026
DOI: 10.5281/zenodo.22858649
Counts differ because each database indexes a different set of publications. We treat OpenAlex as the canonical count; Google Scholar is not shown (no API, and crawling it violates its ToS).
Multi-agent AI systems can retain shared histories, delegation patterns, review procedures, and source reputations across tasks. This paper asks whether agents that respond correctly to matched corrective evidence at the local level can nevertheless participate in collectives that resist the same correction. We define collective governability as the capacity to route material discrepancies into consequential review and to revise binding commitments and the relational configuration through which they are formed. The framework separates the current commitment from correction-relevant relational state, distinguishes commitment revision from relational and constitutional revision, and treats structural correction as a causal claim that must improve performance on matched held-out challenges rather than merely change organizational form. The experimental program uses five modules: local object-level and procedural correctability; history and review routing; self-specific meta-authority; succession; and collective self-revision of a review rule. Correction resistance is assessed against an explicit decision benchmark and, for valid versus invalid challenges, by separating discrimination from a criterion shift toward preserving the status quo. The design uses two frozen decision points, aggregation-matched independent controls, equal-information review-path interventions, designated neutral review stewards, domain-learning and non-diagnostic controls for meta-authority, and succession tests that isolate public status and historical linkage from mechanical permissions and memory-mediated advantages. The paper does not test spontaneous formation of correction-resistant institutions and reports no empirical results. Its contribution is a controlled experimental decomposition of when correction-relevant relational conditions become an additional object of multi-agent safety evaluation beyond the current responses of individual agents.
No comments yet — start the discussion below.