# Bayesian parameter estimation and adaptive control of Markov processes with time-averaged cost

Applicationes Mathematicae (1998)

- Volume: 25, Issue: 3, page 339-358
- ISSN: 1233-7234

This paper considers Bayesian parameter estimation and an associated adaptive control scheme for controlled Markov chains and diffusions with time-averaged cost. Asymptotic behaviour of the posterior law of the parameter given the observed trajectory is analyzed. This analysis suggests a "cost-biased" estimation scheme and associated self-tuning adaptive control. This is shown to be asymptotically optimal in the almost sure sense.

