资源论文Speeding Up Exact Solutions of Interactive Dynamic Influence Diagrams Using Action Equivalence

Speeding Up Exact Solutions of Interactive Dynamic Influence Diagrams Using Action Equivalence

2019-11-15 | |  66 |   39 |   0

Abstract Interactive dynamic inflfluence diagrams (I-DIDs) are graphical models for sequential decision making in partially observable settings shared by other agents. Algorithms for solving I-DIDs face the challenge of an exponentially growing space of candidate models ascribed to other agents, over time. Previous approach for exactly solving IDIDs groups together models having similar solutions into behaviorally equivalent classes and updates these classes. We present a new method that, in addition to aggregating behaviorally equivalent models, further groups models that prescribe identical actions at a single time step. We show how to update these augmented classes and prove that our method is exact. The new approach enables us to bound the aggregated model space by the cardinality of other agents’ actions. We evaluate its performance and provide empirical results in support

上一篇:Efficient Computation of Jointree Bounds for Systematic MAP Search

下一篇:A General Approach to Environment Design with One Agent

用户评价
全部评价

热门资源

  • Stratified Strate...

    In this paper we introduce Stratified Strategy ...

  • The Variational S...

    Unlike traditional images which do not offer in...

  • Learning to learn...

    The move from hand-designed features to learned...

  • A Mathematical Mo...

    Direct democracy, where each voter casts one vo...

  • Learning to Predi...

    Much of model-based reinforcement learning invo...