A Direct Approximation of AIXI Using Logical State Abstractions
Samuel Yang-Zhao, Tianyu Wang, Kee Siong Ng
Abstract
We propose a practical integration of logical state abstraction with AIXI, a Bayesian optimality notion for reinforcement learning agents, to significantly expand the model class that AIXI agents can be approximated over to complex history-dependent and structured environments. The state representation and reasoning framework is based on higher-order logic, which can be used to define and enumerate complex features on non-Markovian and structured environments. We address the problem of selecting the right subset of features to form state abstractions by adapting the -MDP optimisation criterion from state abstraction theory. Exact Bayesian model learning is then achieved using a suitable generalisation of Context Tree Weighting over abstract state sequences. The resultant architecture can be integrated with different planning algorithms. Experimental results on controlling epidemics on large-scale contact networks validates the agent's performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Dynamic Knowledge Injection for AIXI AgentsSamuel Yang-Zhao, Kee Siong Ng, Marcus HutterAAAI 2024
- Self-Predictive Universal AIElliot Catt, Jordi Grau-Moya, Marcus Hutter, Matthew Aitchison et al.NeurIPS 2023
Builds on1
Related papers
- Evolving AgentsLeonardo RanaldiACL 2026 · 227 citations
- Integrating Suboptimal Human Knowledge with Hierarchical Reinforcement Learning for Large-Scale Multiagent SystemsDingbang Liu, Shohei Kato, Wen Gu, Fenghui Ren et al.NeurIPS 2024 · 2 citations
- Planning to the Information Horizon of BAMDPs via Epistemic State AbstractionDilip Arumugam, Satinder SinghNeurIPS 2022 · 7 citations
- Predictive Coding Enhances Meta-RL To Achieve Interpretable Bayes-Optimal Belief Representation Under Partial ObservabilityPo-Chen Kuo, Han Hou, Will Dabney, Edgar Y. WalkerNeurIPS 2025
- Planning with Abstract Learned Models While Learning Transferable SubtasksJohn Winder, Stephanie Milani, Matthew Landen, Erebus Oh et al.AAAI 2020 · 10 citations
