ppo
22 个项目 · ⭐ 65.0kAn elegant PyTorch deep reinforcement learning library.
Simple Reinforcement learning tutorials, 莫烦Python 中文AI教学
Learn Deep Reinforcement Learning in 60 days! Lectures & Code in Python. Reinforcement Learning + Deep Learning
PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Implementations from the free course Deep Reinforcement Learning with Tensorflow and PyTorch
Implementations of basic RL algorithms with minimal lines of codes! (pytorch based)
Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch
Proximal Policy Optimization (PPO) algorithm for Super Mario Bros
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
Long-Term Evolution Project of Reinforcement Learning
chatglm-6b微调/LORA/PPO/推理, 样本为自动生成的整数/小数加减乘除运算, 可gpu/cpu
共 22 条 · 第 1 / 2 页