Open Issues Need Help
View All on GitHub Bug: PPO training diverges on continuous action spaces about 11 hours ago
bug enhancement help wanted good first issue question
Reinforcement learning on AMD GPUs — PPO, SAC, custom envs, RLHF
Python
#open-source#ppo#python#rocm#sac