Mostrar el registro sencillo del ítem
Absorbing Markov Decision Processes
hal.structure.identifier | Méthodes avancées d’apprentissage statistique et de contrôle [ASTRAL] | |
hal.structure.identifier | Institut de Mathématiques de Bordeaux [IMB] | |
hal.structure.identifier | Institut Polytechnique de Bordeaux [Bordeaux INP] | |
dc.contributor.author | DUFOUR, François | |
dc.contributor.author | PRIETO-RUMEAU, Tomás | |
dc.date | 2024 | |
dc.date.accessioned | 2024-04-04T02:31:31Z | |
dc.date.available | 2024-04-04T02:31:31Z | |
dc.date.issued | 2024 | |
dc.identifier.issn | 1292-8119 | |
dc.identifier.uri | https://oskar-bordeaux.fr/handle/20.500.12278/190315 | |
dc.description.abstractEn | In this paper, we study discrete-time absorbing Markov Decision Processes (MDP) with measurable state space and Borel action space with a given initial distribution.For such models, solutions to the characteristic equation that are not occupation measures may exist. Several necessary and sufficient conditions are provided to guarantee that any solution to the characteristic equation is an occupation measure.Under the so-called continuity-compactness conditions, we first show that a measure is precisely an occupation measure if and only if it satisfies the characteristic equation and an additional absolute continuity condition. Secondly, it is shown that the set of occupation measures is compact in the weak-strong topology if and only if the model is uniformly absorbing. Several examples are provided to illustrate our results. | |
dc.language.iso | en | |
dc.publisher | EDP Sciences | |
dc.rights.uri | http://creativecommons.org/licenses/by/ | |
dc.title.en | Absorbing Markov Decision Processes | |
dc.type | Article de revue | |
dc.subject.hal | Mathématiques [math] | |
bordeaux.journal | ESAIM: Control, Optimisation and Calculus of Variations | |
bordeaux.hal.laboratories | Institut de Mathématiques de Bordeaux (IMB) - UMR 5251 | * |
bordeaux.institution | Université de Bordeaux | |
bordeaux.institution | Bordeaux INP | |
bordeaux.institution | CNRS | |
bordeaux.peerReviewed | oui | |
hal.identifier | hal-04377071 | |
hal.version | 1 | |
hal.popular | non | |
hal.audience | Internationale | |
hal.origin.link | https://hal.archives-ouvertes.fr//hal-04377071v1 | |
bordeaux.COinS | ctx_ver=Z39.88-2004&rft_val_fmt=info:ofi/fmt:kev:mtx:journal&rft.jtitle=ESAIM:%20Control,%20Optimisation%20and%20Calculus%20of%20Variations&rft.date=2024&rft.eissn=1292-8119&rft.issn=1292-8119&rft.au=DUFOUR,%20Fran%C3%A7ois&PRIETO-RUMEAU,%20Tom%C3%A1s&rft.genre=article |
Archivos en el ítem
Archivos | Tamaño | Formato | Ver |
---|---|---|---|
No hay archivos asociados a este ítem. |