I worked on developing a reinforcement learning model using TensorFlow. From research, policy-gradient based reinforcement was one of the best machine learning algorithms to tackle this sort of time-varying 3D space problem. There was not enough time to train it and so a different approach was implemented in the submitted version to at least have a fairly completed project.
Log in or sign up for Devpost to join the conversation.