Hedging as Reward Augmentation in Probabilistic Graphical Models
Debarun Bhattacharjya, Radu Marinescu
摘要
Most people associate the term ‘hedging’ exclusively with financial applications, particularly the use of financial derivatives. We argue that hedging is an activity that human and machine agents should engage in more broadly, even when the agent’s value is not necessarily in monetary units. In this paper, we propose a decision-theoretic view of hedging based on augmenting a probabilistic graphical model – specifically a Bayesian network or an influence diagram – with a reward. Hedging is therefore posed as a particular kind of graph manipulation, and can be viewed as analogous to control/intervention and information gathering related analysis. Effective hedging occurs when a risk-averse agent finds opportunity to balance uncertain rewards in their current situation. We illustrate the concepts with examples and counter-examples, and conduct experiments to demonstrate the properties and applicability of the proposed computational tools that enable agents to proactively identify potential hedging opportunities in real-world situations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Dynamic Knowledge Injection for AIXI AgentsSamuel Yang-Zhao, Kee Siong Ng, Marcus HutterAAAI 2024
- Rehearsal Learning for Avoiding Undesired FutureTian Qin, Tian-Zuo Wang, Zhi-Hua ZhouNeurIPS 2023 · 被引用 8 次
- Agent Incentives: A Causal PerspectiveTom Everitt, Ryan Carey, Eric D. Langlois, Pedro A. Ortega 等AAAI 2021 · 被引用 66 次
- A New Bounding Scheme for Influence DiagramsRadu Marinescu, Junkyu Lee, Rina DechterAAAI 2021 · 被引用 1 次
- How About Kind of Generating Hedges using End-to-End Neural Models?Alafate Abulimiti, Chloé Clavel, Justine CassellACL 2023 · 被引用 2 次
