TY - GEN
T1 - Optimal value of information in dynamic Bayesian networks
AU - Ghosh, Sarthak
AU - Ramakrishnan, C. R.
N1 - Publisher Copyright:
© 2017 IEEE.
PY - 2017/7/2
Y1 - 2017/7/2
N2 - Decision-making based on probabilistic reasoning often involves selecting a subset of expensive observations, that best predict the system state. Krause and Guestrin described two problems of non-myopically selecting observations in graphical models to optimize the value of information (VoI), namely, selection of an optimal subset of observations, and generation of an optimal conditional observation plan. They showed that these problems are intractable in general, but gave polynomial-Time dynamic programming algorithms, called VoIDP, for chain graphical models. In this paper, we consider the general setting of Dynamic Bayesian Networks (DBNs), and formulate these problems in terms of finding optimal policies in Markov Decision Processes (MDPs). The time complexities of the resulting algorithms are exponential in general, but polynomial for chain models. Given a chain model, our algorithms compute the same subset, or plan, as VoIDP. Interestingly, despite their generality, our algorithms have significantly better time complexities for chain models compared to VoIDP. We also present an outline of how to use our framework to formulate an approximate, nonmyopic VoI optimization technique, with absolute a posteriori guarantees on approximation, that can handle arbitrary DBNs efficiently.
AB - Decision-making based on probabilistic reasoning often involves selecting a subset of expensive observations, that best predict the system state. Krause and Guestrin described two problems of non-myopically selecting observations in graphical models to optimize the value of information (VoI), namely, selection of an optimal subset of observations, and generation of an optimal conditional observation plan. They showed that these problems are intractable in general, but gave polynomial-Time dynamic programming algorithms, called VoIDP, for chain graphical models. In this paper, we consider the general setting of Dynamic Bayesian Networks (DBNs), and formulate these problems in terms of finding optimal policies in Markov Decision Processes (MDPs). The time complexities of the resulting algorithms are exponential in general, but polynomial for chain models. Given a chain model, our algorithms compute the same subset, or plan, as VoIDP. Interestingly, despite their generality, our algorithms have significantly better time complexities for chain models compared to VoIDP. We also present an outline of how to use our framework to formulate an approximate, nonmyopic VoI optimization technique, with absolute a posteriori guarantees on approximation, that can handle arbitrary DBNs efficiently.
KW - Dynamic Bayesian Networks
KW - Markov Decision Processes
KW - Optimization algorithms
KW - Probabilistic reasoning
KW - Reasoning under uncertainty
KW - Value of information
UR - https://www.scopus.com/pages/publications/85048513340
U2 - 10.1109/ICTAI.2017.00015
DO - 10.1109/ICTAI.2017.00015
M3 - Conference contribution
AN - SCOPUS:85048513340
T3 - Proceedings - International Conference on Tools with Artificial Intelligence, ICTAI
SP - 16
EP - 23
BT - Proceedings - 2017 International Conference on Tools with Artificial Intelligence, ICTAI 2017
PB - IEEE Computer Society
T2 - 29th IEEE International Conference on Tools with Artificial Intelligence, ICTAI 2017
Y2 - 6 November 2017 through 8 November 2017
ER -