Skip to content
View woithook's full-sized avatar
🏠
Working from home
🏠
Working from home

Highlights

  • Pro

Block or report woithook

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Popular repositories Loading

  1. LIO LIO Public

    A pytorch reproduction of LIO(Learning to Incentivize Other)

    Python 1

  2. A2C-Pytorch-implementations A2C-Pytorch-implementations Public

    Implement the A2C(Advantage Actor-Critic) algorithm using pytorch in multiple environments of openai gym. (Including Cartpole, LunarLander, Pong. Breakout is tuning and maybe complete soon.) Someti…

    Python 1 1

  3. Meta-gradient_RL Meta-gradient_RL Public

    A toy implementation of paper "Meta-Gradient Reinforcement Learning"

    Jupyter Notebook 1

  4. llm-mechanisms-lab llm-mechanisms-lab Public

    Personal notes documenting my exploration, implementation, and research on Large Language Models (LLMs).

    Jupyter Notebook 1

  5. ray_exercise ray_exercise Public

    Some exercises performed during learning ray/tune/rllib

    Python

  6. gym gym Public

    Forked from openai/gym

    A toolkit for developing and comparing reinforcement learning algorithms.

    Python