Learning the structure of Factored Markov Decision Processes in reinforcement learning problems
Thomas Degris(Google DeepMind (United Kingdom)), Pierre-Henri Wuillemin(Centre National de la Recherche Scientifique), Olivier Sigaud
Cited by 90
Related Papers
Deterministic policy gradient algorithms
|HAL (Le Centre pour la Communication Scientifique Directe)|2014|1.7k
Insulin resistance and inflammation predict kinetic body weight changes in response to dietary weight loss and maintenance in overweight and obese subjects by using a Bayesian network approach
|American Journal of Clinical Nutrition|2013|86
From Motor Learning to Interaction Learning in Robots
|Studies in computational intelligence|2009|73
Towards a global modelling of the Camembert-type cheese ripening process by coupling heterogeneous knowledge with dynamic Bayesian networks
|Journal of Food Engineering|2009|44
Top-Down Construction and Repetetive Structures Representation in Bayesian Networks
|The Florida AI Research Society|2000|42