
רונן ברפמן
אקדמי בכיר
Regular decision processes
Modelling dynamic systems without using hidden variables
We describe Regular Decision Processes (RDPs) a model in between MDPs and POMDPs. Like in POMDPs, the effect of an action may depend on the entire history of actions and observations, but this dependence is restricted to regular functions only. This makes RDP a tractable, yet rich model, that does not hypothesize hidden state, and could possibly be useful for learning dynamic systems.
| שפת פרסום | אנגלית |
| דפים | 1844-1846 |
| סטטוס פרסום | פורסם - 01.01.2019 |
ASJC Scopus subject areas
Artificial Intelligence
Software
Control and Systems Engineering