Omega-Regular Decision Processes

Hahn, Ernst Moritz; Perez, Mateo; Schewe, Sven; Somenzi, Fabio; Trivedi, Ashutosh; Wojtczak, Dominik

Full-text links:

Download:

Current browse context:

cs.LO

< prev | next >

new | recent | 2312

Computer Science > Logic in Computer Science

Title: Omega-Regular Decision Processes

Authors: Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi, Ashutosh Trivedi, Dominik Wojtczak

(Submitted on 14 Dec 2023)

Abstract: Regular decision processes (RDPs) are a subclass of non-Markovian decision processes where the transition and reward functions are guarded by some regular property of the past (a lookback). While RDPs enable intuitive and succinct representation of non-Markovian decision processes, their expressive power coincides with finite-state Markov decision processes (MDPs). We introduce omega-regular decision processes (ODPs) where the non-Markovian aspect of the transition and reward functions are extended to an omega-regular lookahead over the system evolution. Semantically, these lookaheads can be considered as promises made by the decision maker or the learning agent about her future behavior. In particular, we assume that, if the promised lookaheads are not met, then the payoff to the decision maker is $\bot$ (least desirable payoff), overriding any rewards collected by the decision maker. We enable optimization and learning for ODPs under the discounted-reward objective by reducing them to lexicographic optimization and learning over finite MDPs. We present experimental results demonstrating the effectiveness of the proposed reduction.

Subjects:	Logic in Computer Science (cs.LO); Machine Learning (cs.LG)
Cite as:	arXiv:2312.08602 [cs.LO]
	(or arXiv:2312.08602v1 [cs.LO] for this version)

Submission history

From: Mateo Perez [view email]
[v1] Thu, 14 Dec 2023 01:58:51 GMT (172kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2312.08602

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Logic in Computer Science

Title: Omega-Regular Decision Processes

Submission history