Pith. sign in

REVIEW 4 cited by

AI Research Considerations for Human Existential Safety (ARCHES)

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2006.04948 v1 pith:ZWBHLCNV submitted 2020-05-30 cs.CY cs.AIcs.LG

classification cs.CYcs.AIcs.LG
keywords researchexistentialmightriskssafetybenefitcontemporarydirection
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Framed in positive terms, this report examines how technical AI research might be steered in a manner that is more attentive to humanity's long-term prospects for survival as a species. In negative terms, we ask what existential risks humanity might face from AI development in the next century, and by what principles contemporary technical research might be directed to address those risks. A key property of hypothetical AI technologies is introduced, called \emph{prepotence}, which is useful for delineating a variety of potential existential risks from artificial intelligence, even as AI paradigms might shift. A set of \auxref{dirtot} contemporary research \directions are then examined for their potential benefit to existential safety. Each research direction is explained with a scenario-driven motivation, and examples of existing work from which to build. The research directions present their own risks and benefits to society that could occur at various scales of impact, and in particular are not guaranteed to benefit existential safety if major developments in them are deployed without adequate forethought and oversight. As such, each direction is accompanied by a consideration of potentially negative side effects.

Discussion (0). Sign in to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 25 citations worldwide. Full citation record

  1. A dataset of rated conceptual arguments

    cs.AI 2026-07 conditional novelty 6.5 of 10

    A multi-dimensional expert-rated dataset of 951 conceptual-argument critiques shows LLM judge performance tracks general model capability and is little helped by reasoning modes.

  2. Information Limits and Attractor Dynamics in Economies of Frontier LLM Agents: A Pre-Registered Test

    cs.AI 2026-07 conditional novelty 6.0 of 10

    A pre-registered experiment on Claude Opus 4.8 agents confirms an information-theoretic capacity region for coupled LLM-agent markets (with one key result being an algebraic identity) and finds that LLM populations ex...

  3. Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets

    cs.AI 2025-05 conditional novelty 5.0 of 10

    AI agents in future labor markets will need metacognitive and strategic reasoning because incomplete information creates adverse selection, moral hazard, and reputation effects.

  4. The Theory of Strategic Evolution: Games with Endogenous Players and Strategic Replicators

    cs.GT 2025-12 reject novelty 4.0 of 10

    A theory of strategic evolution says multi-level systems of self-reproducing optimizers are stable only under a small-gain condition, and stable AI alignment requires bounding self-modification.

Pith tools