arXiv · 2110.02450
Reward-Punishment Symmetric Universal Intelligence
Abstract
Can an agent's intelligence level be negative? We extend the Legg-Hutter agent-environment framework to include punishments and argue for an affirmative answer to that question. We show that if the background encodings and Universal Turing Machine (UTM) admit certain Kolmogorov complexity symmetries, then the resulting Legg-Hutter intelligence measure is symmetric about the origin. In particular, this implies reward-ignoring agents have Legg-Hutter intelligence 0 according to such UTMs.
Explore related subjects
Keep this discovery
Samuel Allen Alexander, Marcus Hutter. 2021-10-06. Reward-Punishment Symmetric Universal Intelligence. https://doi.org/10.1007/978-3-030-93758-4_1
Cite the original work for its findings. Save a collection to share your selection of sources.