Author
Omar Rivasplata
0 works0 citations
Recent research
- AI & ComputingOpen access
Semi-pessimistic Reinforcement Learning
Offline reinforcement learning aims to learn an optimal policy from pre-collected data. However, it faces challenges of distributional shift, where the learned policy may encounter unseen scenarios not covered in the offline data. Additionally, numerous applications suffer from a...