Delighted to share our new simulation solution for parallel reinforcement learning. This allows us to train up to 74 agents simultaneously and reduces training time from 3.9 hours to 11 minutes. The paper and code can be found here:
arxiv.org/abs/2209.11094
github.com/SaundersJE97/P…
An Associate Professor in Robotics at @UniofBath



