rl
31 个项目 · ⭐ 73.8kLlama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用
An elegant PyTorch deep reinforcement learning library.
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
A Survey of Reinforcement Learning for Large Reasoning Models
Implementation of all RL algorithms in a simpler way
Scalable RL solution for advanced reasoning of language models
The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
Tensorlake is a serverless runtime for sandboxes and deploying background agentic applications
A universal flight control tuning framework
Agentic RAG R1 Framework via Reinforcement Learning
共 31 条 · 第 1 / 2 页