vix.ing · top · new · best · stats · spec

Partially Observable Discrete-time Discounted Markov Games with General Utility

2022/11/15 by Arnab Bhabak, Bhabak, Arnab, Subhamay Saha +1 · 1 citation
Decision Sciences · Economics, Econometrics and Finance · #Economic theories and models #FOS: Mathematics #Game Theory and Applications #Optimization and Control (math.OC)

paper · pdf · doi:10.48550/arxiv.2211.07888

openalex publication_date 2022/11/15 · openalex created_date 2022/11/24 · openalex updated_date 2026/07/28

Abstract

In this paper, we investigate a partially observable zero sum games where the state process is a discrete time Markov chain. We consider a general utility function in the optimization criterion. We show the existence of value for both finite and infinite horizon games and also establish the existence of optimal polices. The main step involves converting the partially observable game into a completely observable game which also keeps track of the total discounted accumulated reward/cost.

Cited by

Related