Hi Lars, thanks for open-sourcing this work!
I’m Yufei Jia, the author of UniLab, which you cited when discussing massively parallel robot RL workloads. We are also working on scalable robot RL infrastructure.
I’m curious whether rl-triton could help accelerate off-policy algorithms such as SAC or TD3. If so, which part of their training pipeline do you think would benefit most?
I’d be happy to exchange notes, especially around massively parallel robot RL and AMD/MI300X.
Great work!
Hi Lars, thanks for open-sourcing this work!
I’m Yufei Jia, the author of UniLab, which you cited when discussing massively parallel robot RL workloads. We are also working on scalable robot RL infrastructure.
I’m curious whether rl-triton could help accelerate off-policy algorithms such as SAC or TD3. If so, which part of their training pipeline do you think would benefit most?
I’d be happy to exchange notes, especially around massively parallel robot RL and AMD/MI300X.
Great work!