Improve TD3 RL Agent Convergence
Budget: £20 – £250 GBP
I'm seeking an expert to optimize the Twin Delayed DDPG (TD3) reinforcement learning agent for a Virtual 2D robotics control task. I can provide working code. The environment is based on OpenAI's bipedalwalker Environment. The goal is to reduce the number of episodes needed for convergence and enhance stability during training.
Key Requirements:
- Leverage existing extensive data to improve convergence.
- Focus on stability and fewer episodes for the TD3 agent.
-Very Fast Completion (within 24 hours)
Ideal Skills and Experience:
- Proven background in reinforcement learning, specifically with TD3.
- Strong understanding of virtual robotics control environments.
- Experience with data-driven optimization techniques.
- Ability to deliver measurable improvements in training efficiency.
Key Requirements:
- Leverage existing extensive data to improve convergence.
- Focus on stability and fewer episodes for the TD3 agent.
-Very Fast Completion (within 24 hours)
Ideal Skills and Experience:
- Proven background in reinforcement learning, specifically with TD3.
- Strong understanding of virtual robotics control environments.
- Experience with data-driven optimization techniques.
- Ability to deliver measurable improvements in training efficiency.