NO

N.A. Ordonez Cardenas

info

Please Note

2 records found

Evaluating state-of-the-art reinforcement learning technique for adaptability to human collaborators

Bachelor thesis (2022) - N.A. Ordonez Cardenas, F.A. Oliehoek, R.T. Loftin
A longstanding problem in the area of reinforcement learning is human-agent col- laboration. As past research indicates that RL agents undergo a distributional shift when they start collaborating with human beings, the goal is to create agents that can adapt. We build upon research using the two-player Overcooked environment to repro- duce a simplified version of the Fictitious Co-Play algorithm in order to confirm past found improvements at a smaller scale of training and using Self-Play and Population- based trained algorithms as the baselines for comparison. We find that the agent on average slightly outperforms both baseline algorithms when evaluated using a human proxy. We also find high cross-seed variance in performance, indicating the potential for further hyperparameter tuning. ...