arXiv · 2505.17882
Formalizing Embeddedness Failures in Universal Artificial Intelligence
Abstract
We rigorously discuss the commonly asserted failures of the AIXI reinforcement learning agent as a model of embedded agency. We attempt to formalize these failure modes and prove that they occur within the framework of universal artificial intelligence, focusing on a variant of AIXI that models the joint action/percept history as drawn from the universal distribution. We also evaluate the progress that has been made towards a successful theory of embedded agency based on variants of the AIXI agent.
Explore related subjects
Keep this discovery
Cole Wyeth, Marcus Hutter. 2025-05-23. Formalizing Embeddedness Failures in Universal Artificial Intelligence. https://arxiv.org/abs/2505.17882
Cite the original work for its findings. Save a collection to share your selection of sources.