Adaptive control of diffusion processes with a discounted reward criterion

Autorzy

Dane publikacji

  • DOI: 10.4064/am2421-10-2020

  • Tom 47

  • Zeszyt 2

  • Czasopismo: Applicationes Mathematicae

  • Strony: 225-253

  • Data publikacji online: 08.12.2020

Liczba wyświetleń: 0

Liczba pobrań: 0

Wersja elektroniczna

Otwarty dostęp

Abstrakt

The optimal control problem we are dealing with in this paper is to determine control policies that maximize a discounted reward criterion when the dynamic system evolves as a stochastic differential equation (SDE). Both the instantaneous reward function and the SDE’s drift coefficient may depend on an unknown parameter. We give conditions ensuring the existence of an asymptotically optimal policy using the so-called Principle of Estimation and Control. We illustrate our results with several examples.