PPO, DDPG, SAC implementation on mujoco environment
-
Updated
Feb 16, 2022 - Python
PPO, DDPG, SAC implementation on mujoco environment
MERL: Multi-Head Reinforcement Learning (TensorFlow).
some experiments with training and fine-tuning decision transformer
Diffusion Policy复现 & Visual Diffusion Policy | 具身智能/模仿学习 | HalfCheetah MuJoCo, PyTorch, 3阶段: CNN基础→状态DP→视觉DP, RTX 4060
7 个经典强化学习算法的 PyTorch 从零实现(DQN / Double / Dueling / PER / REINFORCE / Actor-Critic / PPO),含中文 README 与可复现训练结果
Five-person offline RL course project: IQL on HalfCheetah-v5 with Minari, shared contracts and reproducible team workflow
To associate your repository with the halfcheetah topic, visit your repo's landing page and select "manage topics."