Pular para o conteúdo principal

ETHNOS_APP

Início • Busca • Periódicos • Lista 0

Artificial Intelligence

Arguments for Catastrophic Risk

Dados Bibliográficos

ID4079996
AutoresAdam Bales (0000-0002-9629-0318, University of Oxford), William D’alessandro (0000-0002-5451-079X, University of Oxford), Cameron Domenico Kirk‐giannini (0000-0001-9372-227X, Rutgers, the State University of New Jersey, autor correspondente)
Ano2024
Volume19
Fascículo2
Data de publicação2024-02-01
Peer ReviewedSim
Open AccessSim
TipoARTICLE
PeriódicoPhilosophy Compass (JOURNAL)
Identificadores do periódicoISSN: 1747-9991 • E-ISSN: 1747-9991
EditoraWiley (PUBLISHER • GB)
DOI10.1111/phc3.12964
OpenAlexW4391719496
IdiomaEN
Citações recebidas14
Referências citadas24

Recent progress in artificial intelligence (AI) has drawn attention to the technology's transformative potential, including what some see as its prospects for causing large-scale harm. We review two influential arguments purporting to show how AI could pose catastrophic risks. The first argument - theProblem of Power-Seeking- claims that, under certain assumptions, advanced AI systems are likely to engage in dangerous power-seeking behavior in pursuit of their goals. We review reasons for thinking that AI systems might seek power, that they might obtain it, that this could lead to catastrophe, and that we might build and deploy such systems anyway. The second argument claims that the development of human-level AI will unlock rapid further progress, culminating in AI systems far more capable than any human - this is theSingularity Hypothesis. Power-seeking behavior on the part of such systems might be particularly dangerous. We discuss a variety of objections to both arguments and conclude by assessing the state of the debate

Argument (complex analysis · Artificial general intelligence · Cognitive science · Epistemology · Harm · Power (physics · Transformative learning · Variety (cybernetics · Computer Science · Ethics and Social Impacts of AI · Neuroethics, Human Enhancement, Biomedical Innovations · Philosophy · Psychology · Social Psychology · Space Science and Extraterrestrial Life · Artificial Intelligence

  • Fear, power, and superintelligence

    Open Access•Carlo Burelli, Federico Formentini•European Journal of Political…•2026

  • Will AI and humanity go to war

    Open Access•Simon Goldstein•AI & Society•2026

  • Is Alignment Unsafe

    Open Access•Cameron Domenico Kirk‐giannini•Philosophy & Technology•2024

  • Language Agents and Malevolent Design

    Open Access•Inchul Yum•Philosophy & Technology•2024

  • ‘Systematic Alignment Decay’

    Open Access•Brendan Kelters•AI & Society•2026

  • Disagreement, AI alignment, and bargaining

    Open Access•Harry R Lloyd•Philosophical Studies•2025

  • Promotionalism, orthogonality, and instrumental convergence

    Open Access•Nathaniel Sharadin•Philosophical Studies•2025

  • Deception and manipulation in generative AI

    Open Access•Christian Tarsney•Philosophical Studies•2025

  • What is AI safety? What do we want it to be

    Open Access•Jacqueline Harding, Cameron Domenico Kirk‐giannini•Philosophical Studies•2025

  • AI safety

    Open Access•Herman Cappelen, Josh Dever et al.•Philosophical Studies•2025

  • Against the singularity hypothesis

    Open Access•David Thorstad•Philosophical Studies•2025

  • Existentialist risk and value misalignment

    Open Access•Ariela Tubert, Justin Tiehen•Philosophical Studies•2025

  • Two types of AI existential risk

    Open Access•Atoosa Kasirzadeh•Philosophical Studies•2025

  • Artificial Intelligence

    Open Access•William D’alessandro, William D'Alessandro et al.•Philosophy Compass•2025

  • Mind Children

    Hans Moravec•Mind Children•1990

  • Highly accurate protein structure prediction with AlphaFold

    Open Access•John Jumper, Richard Evans et al.•Nature•2021

  • Artificial Intelligence, Values, and Alignment

    Open Access•Ingeborg Gabriel•Minds and Machines•2020

  • Existential risk from AI and orthogonality

    Open Access•Vincent C Müller, Michael Cannon et al.•Ratio•2022

  • Language agents reduce the risk of existential catastrophe

    Open Access•Simon Goldstein, Cameron Domenico Kirk‐giannini•AI & Society•2025

  • Will AI avoid exploitation? Artificial general intelligence and expected utility theory

    Open Access•Adam Bales•Philosophical Studies•2025

Obras citantes distintas14
Citações por ano7
Intervalo de citações2024 - 2026 (3)
Velocidade de citaçãocurrent
Altamente citadoNão
Tipos de citaçãoNeutras: 13
Ethnos_APP • Projeto Open Source • Licença MIT • Frontend v2.0.0 • Privacidade e Cookies • Documentação da API: api.ethnos.app/docs • Código da API: GitHub • DOI: 10.5281/zenodo.17049435 • Código do Frontend: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae