Reinforcement Learning
I am exploring hybrid policy improvement methods that combine on-policy and off-policy learning.
- My current focus is on how these approaches can complement each other within a policy improvement framework.
- I am reviewing related work and developing the research direction, with robot learning as a potential application.