reinforcement-learning
233 个项目 · ⭐ 1118.6kPython library for solving reinforcement learning (RL) problems using generative models (e.g. Diffusion Models).
A framework for agentic tool use training with reinforcement learning
Notebooks for the Practicals at the Deep Learning Indaba 2022.
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
Stable and Efficient Reinforcement Learning for Trillion-Parameter LLMs
JAX implementation of WSRL and RL baselines | ICLR 2025
Seed, Code, Harvest: Grow Your Own App with Tree of Thoughts!
303 份 AI/LLM 中文讲义,支持在线阅读、PDF 下载和 LaTeX 源码查看 | Stanford CS336/CS224R/CS25 | Berkeley LLM Agents | Agent 工程实践
Your AI intranet: network the computers you already own for inference and training.
NEWTON: Agentic Planning for Physically Grounded Video Generation
Efficient World Models with Context-Aware Tokenization. ICML 2024
An Open-Source Package for Chinese Open-domain Conversational Chatbot (中文闲聊对话系统,一键部署微信闲聊机器人)
共 233 条 · 第 9 / 12 页