Entropy-Regularized Partially Observed Markov Decision Processes

2021-12-22Unverified0· sign in to hype

Timothy L. Molloy, Girish N. Nair

Unverified — Be the first to reproduce this paper.

Abstract

We investigate partially observed Markov decision processes (POMDPs) with cost functions regularized by entropy terms describing state, observation, and control uncertainty. Standard POMDP techniques are shown to offer bounded-error solutions to these entropy-regularized POMDPs, with exact solutions possible when the regularization involves the joint entropy of the state, observation, and control trajectories. Our joint-entropy result is particularly surprising since it constitutes a novel, tractable formulation of active state estimation.

Tasks

State Estimation

Entropy-Regularized Partially Observed Markov Decision Processes

Abstract

Tasks

Reproductions