Skip to main content

Envisioning is an emerging technology research institute and advisory.

LinkedInInstagramGitHub

2011 — 2026

research
  • Observatory
  • Newsletter
  • Methodology
  • Origins
  • Vocab
  • RSS Feeds
services
  • Signals Session
  • Bespoke Projects
  • Build Sessions
  • Pricing
  • Use Cases
  • Signals
  • Signal Scan↗free
impact
  • ANBIMAFuture of Brazilian Capital Markets
  • IEEECharting the Energy Transition
  • Horizon 2045Future of Human and Planetary Security
  • WKOTechnology Scanning for Austria
solutions
  • Innovation
  • Strategy
  • Consultants
  • Foresight
  • Associations
  • Governments
  • L&D
resources
  • Partners
  • Coding for Non-Coders
  • How We Work
  • Data Visualization
  • Multi-Model Method
  • FAQ
  • Security & Privacy
  • Public Sector
about
  • Manifesto
  • Community
  • Events
  • Support
  • Contact
ResearchServicesSignalsAbout
ResearchServicesSignalsAbout
  1. Home
  2. Vocab
  3. Ethics-as-Cooperation-Technology

Ethics-as-Cooperation-Technology

Ethics reframed as the technology by which agents make themselves legible and coordinate.

Year: 2025Generality: 450Added: Jul 12, 2026
Back to Vocab

Ethics-as-cooperation-technology is the view that ethical norms function as multiplayer coordination technologies: shared rules, conventions, and heuristics that allow multiple agents to make themselves legible to one another, sustain cooperation across populations with heterogeneous information, and keep cooperative equilibria from collapsing when conditions change. Rather than treating ethics as a ranking of individual actions against abstract principles, the framing treats it as an engineering discipline focused on the conditions under which stable cooperative arrangements can emerge and persist across populations that include agents of varying capability, motivation, and access to information. The framing is developed in alignment research and draws on evolutionary biology, decision theory, and mechanism design, but its specific articulation as a coordination-technology rather than as ethical theory traces to soft-money research from the lab Softmax in 2025, framed as a complement to single-agent value training in frontier AI labs.

The mechanism maps onto concrete properties that distinguish robust cooperative norms from brittle ones. Useful norms, in this framing, are legible: other agents can predict what you will do. They are teachable: weaker or less informed agents can acquire them. They are predictably generalizable: they transfer to novel scenarios, even by less capable agents, degrading gracefully rather than catastrophically when applied outside their training distribution. They are robust: small deviations in conditions, available compute, or population composition do not produce disastrously misaligned behavior. They are corrigible: they update as the population learns. And they are incentive-compatible: maintaining them does not require heroic self-sacrifice from individual agents. These properties are the engineering specification; ethics, in this view, is the discipline of designing norms that satisfy them.

The tradeoff against single-agent value training — the standard frontier-lab approach of specifying values and instilling them in a frozen model — is consequential. Single-agent training produces a coherent behavior profile in one system but leaves the population-level coordination problem unsolved. Agents trained this way may each be individually well-behaved yet produce collectively catastrophic dynamics when they interact, because their training environments did not require them to anticipate or coordinate with other capable agents in the first place. Ethics-as-cooperation-technology trades some of the verifiability of single-agent value loading for a population-level reliability that may matter more as AI systems increasingly operate in mixed human-AI populations, encounter situations their training never covered, and begin self-evolving in ways that break the implicit assumption that frozen preferences remain appropriate.

Open questions in the framing include whether the engineering properties above are jointly satisfiable in practice, whether the discipline can be operationalized as a set of training objectives a frontier model could be optimized against, and how the framing relates to existing traditions in moral philosophy beyond its design specifications. A related unresolved issue is whether the population-level lens is sufficient on its own: there remain cases, including those involving non-cooperative populations or asymmetric capabilities, where individual-level ethical reasoning seems necessary in addition to coordination hygiene. The view's proponents treat the framing as additive rather than substitutive, intended to expand the design space for aligned behavior rather than to retire other approaches.

Research this in Signals

Scan Ethics-as-Cooperation-Technology for yourself.

Signals turns a topic into a sourced research record you can inspect and rerun. Your first scan is free, and this one starts with Ethics-as-Cooperation-Technology already loaded, so edit it or scan as is.