Skip to content

Applicability to off-policy RL algorithms #26

Description

@TATP-233

Hi Lars, thanks for open-sourcing this work!

I’m Yufei Jia, the author of UniLab, which you cited when discussing massively parallel robot RL workloads. We are also working on scalable robot RL infrastructure.

I’m curious whether rl-triton could help accelerate off-policy algorithms such as SAC or TD3. If so, which part of their training pipeline do you think would benefit most?

I’d be happy to exchange notes, especially around massively parallel robot RL and AMD/MI300X.

Great work!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    questionFurther information is requested

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions