Controllability as a Core Principle for AGI Governance and Safety

In Manuel F. Silva, Mohammad Osman Tokhi, Maria Isabel A. Ferreira, Benedita Malheiro, Pedro Guedes, Paulo Ferreira & Maria Teresa Costa, Crisis or Redemption with AI and Robotics? The Dawn of a New Era: Proceedings of the ICRES 2025 Conference. Cham: Springer Nature Switzerland. pp. 144-153 (2026)
  Copy   BIBTEX

Abstract

This paper explores the importance of ensuring AGI (Artificial General Intelligence) controllability and safety as AI systems advance from narrow AI (NAI) to more autonomous systems. AGI's ability to learn and make decisions independently introduces significant challenges, particularly in critical sectors like healthcare, finance, infrastructure, and also when it is embedded in a physical object, such as a robot. Traditional AI governance principles of transparency, explainability, and accountability become insufficient when dealing with more sophisticated AGI, which, in this paper, refers to AI systems that exhibits significantly higher autonomy than current systems, as these models are too complex to be fully understood or controlled by humans. Instead, the paper argues that “controllability” element should be the primary focus of AGI governance to prevent unintended consequences.The paper examines technological approaches such as control by design, fail-safes and redundancy mechanisms, formal verification, adversarial testing, adaptive ethical constraints and sandboxing, alongside institutional strategies including business continuity planning, continuous monitoring, AI ethics boards, and multi-layered audits. It stresses that a combination of technological, institutional, and regulatory measures is essential to ensure AGI remains safe and aligned with human intent. The paper concludes by emphasizing the need for interdisciplinary collaboration among engineers, ethicists, legal experts and policymakers and calls for AGI development to be guided by human values and governance frameworks to avoid catastrophic risks and ensure that AI serves societal benefits.*Please note that this research is preliminary and intended to serve as a basis for open discussion during the Special Session.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

Comprehensive Review of Artificial General Intelligence AGI, Agentic AI and GenAI: Current Trends and Future Directions.Satyadhar Joshi - 2025 - International Journal of Multidisciplinary Research and Growth Evaluation 6 (3):681-688.
Towards Ethical Foundation Models in Robotics: Challenges and Proposals.Ergina Kavallieratou - 2026 - In Manuel F. Silva, Mohammad Osman Tokhi, Maria Isabel A. Ferreira, Benedita Malheiro, Pedro Guedes, Paulo Ferreira & Maria Teresa Costa, Crisis or Redemption with AI and Robotics? The Dawn of a New Era: Proceedings of the ICRES 2025 Conference. Cham: Springer Nature Switzerland. pp. 15-21.

Analytics

Added to PP
2025-11-18

Downloads
28 (#1,646,389)

6 months
23 (#449,438)

Historical graph of downloads
How can I increase my downloads?

Citations of this work

No citations found.

Add more citations

References found in this work

No references found.

Add more references