BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//University of Liverpool Computer Science Seminar System//v2//EN
BEGIN:VEVENT
DTSTAMP:20260919T231627Z
UID:Seminar-DMML-524@lxserverM.csc.liv.ac.uk
ORGANIZER:CN=Danushka Bollegala:MAILTO:Danushka.Bollegala@liverpool.ac.uk
DTSTART:20170908T130000
DTEND:20170908T140000
SUMMARY:Data Mining and Machine Learning Series
DESCRIPTION:Frans Oliehoek: Bayesian Reinforcement Learning for Problems with State Uncertainty\n\nSequential decision making under uncertainty is a challenging problem, especially when the decision maker, or agent, has uncertainty about what the true 'state' of the environment is. That is, in many applications the problem is 'partially observable': there are important pieces of information that are fundamentally hidden from the agent. Moreover, the problem gets even more complex when no accurate model of the environment is available. In such cases, the agent will need to update its belief over the environment, i.e., learn, during execution.\n\nIn this talk, I will introduce a formal way of modeling decision making under partial observability, as well as a more recent extension to the learning setting. I will explain how the learning problem can be tackled using a method called 'POMCP', and how this can be made more efficient via a number of novel techniques. Time permitting, I will also discuss extensions of this methodology that explicitly deal with coordination with other agents, and anticipation of other actors (such as humans) in the environment.\n\nhttps://www.csc.liv.ac.uk/research/seminars/abstract.php?id=524
LOCATION:
END:VEVENT
END:VCALENDAR
