Unmonitorability of Artificial Intelligence

Roman Yampolskiy

Unmonitorability of Artificial Intelligence

Abstract

Artificially Intelligent (AI) systems have ushered in a transformative era across various domains, yet their inherent traits of unpredictability, unexplainability, and uncontrollability have given rise to concerns surrounding AI safety. This paper aims to demonstrate the infeasibility of accurately monitoring advanced AI systems to predict the emergence of certain capabilities prior to their manifestation. Through an analysis of the intricacies of AI systems, the boundaries of human comprehension, and the elusive nature of emergent behaviors, we argue for the impossibility of reliably foreseeing some capabilities. By investigating these impossibility results, we shed light on their potential implications for AI safety research and propose potential strategies to overcome these limitations.

View on PhilPapers

Author's Profile

Roman Yampolskiy

University of Louisville

Archival history

Archival date: 2023-06-11
View all versions

Keywords

Monitoring AI Audit Observability AI Safety

Reprint years

Analytics

Added to PP
2023-06-11

Downloads
896 (#26,188)

6 months
101 (#61,461)

Historical graph of downloads since first upload

This graph includes both downloads from PhilArchive and clicks on external links on PhilPapers.

How can I increase my downloads?

Applied ethics	Epistemology	History of Western Philosophy	Meta-ethics	Metaphysics	Normative ethics
Philosophy of biology	Philosophy of language	Philosophy of mind	Philosophy of religion	Science Logic and Mathematics	More ...

Unmonitorability of Artificial Intelligence

Abstract

Author's Profile

Archival history

Categories

Keywords

Reprint years

Analytics