PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 10, 20260 citationsOpen Access

Gold-Standard AGI: What it is, and how to build it (including an implementation-neutral solution to the outer AGI superalignment problem)

View Full Paper
ATAaron Turner

Key Points

  • The aim is to outline a foundational theory for AGI that addresses alignment issues to benefit humanity as a whole.
  • Explored concepts of outer and inner alignment in AGI development.
  • Presented a foundational theory of AGI and its implications.
  • Designed an implementation-neutral solution for the outer AGI alignment problem.
  • Outlined key challenges in defining final goals for AGI that reflect human values.
  • Developed a theoretical framework for superalignment to ensure AGI's goals align with humanity's.
  • Emphasized the importance of accessible information for AGI policymakers.

Abstract

The way in which AI (and, in particular, agentic superintelligent AGI) develops over the coming decades will determine the fate of all humanity for all eternity. In order to maximise the net benefit of AGI for all humanity, without favouring any subset thereof, we imagine a Gold-Standard AGI that is maximally-aligned and maximally-validated. The first of these properties --- alignment --- is traditionally decomposed into outer alignment (how do we define a final goal FGG that correctly states what we want? ), and inner alignment (how do we build an agent G that forever pursues FGG as intended? ) This paper presents a complete and foundational theory of AGI, culminating in an implementation-neutral solution to the outer AGI alignment problem in the case that G is superintelligent (hence "superalignment"). Given the AGI alignment problem's profound relevance to AGI governance, we adopt a pedagogic style throughout in order that the paper might be accessible to less technical readers such as AGI policymakers.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Aaron Turner (2026) studied this question.

synapsesocial.com/papers/69d895ea6c1944d70ce0713ahttps://doi.org/10.5281/zenodo.19476760
Ask AI
Helpful
Bookmark
Share
View Full Paper