On monitorability of AI

Roman V. Yampolskiy

Journal ArticleOPEN ACCESS

On monitorability of AI

Yampolskiy R

AI and Ethics (2024)

DOI: 10.1007/s43681-024-00420-x

N/ACitations

22Readers

Abstract

Artificially intelligent (AI) systems have ushered in a transformative era across various domains, yet their inherent traits of unpredictability, unexplainability, and uncontrollability have given rise to concerns surrounding AI safety. This paper aims to demonstrate the infeasibility of accurately monitoring advanced AI systems to predict the emergence of certain capabilities prior to their manifestation. Through an analysis of the intricacies of AI systems, the boundaries of human comprehension, and the elusive nature of emergent behaviors, we argue for the impossibility of reliably foreseeing some capabilities. By investigating these impossibility results, we shed light on their potential implications for AI safety research and propose potential strategies to overcome these limitations.

Cite

CITATION STYLE

APA

Yampolskiy, R. V. (2024). On monitorability of AI. AI and Ethics. https://doi.org/10.1007/s43681-024-00420-x

On monitorability of AI

Abstract

Cite

Register to see more suggestions