reinforcement-learning
233 个项目 · ⭐ 1118.6k[ICML 2026] a unified reinforcement learning toolbox for joint RL on language models and diffusion models
FlowSteer: agents designing agentic workflows via reinforced progressive canvas editing.
An implementation of GRPO for Unsloth's VLMs training
[ICLR 2025] The offical implementation of "PSEC: Skill Expansion and Composition in Parameter Space", a new framework designed to facilitate efficient and flexible skill expansion and composition, iteratively evolve the agents' capabilities and efficiently address new challenges
ADvISER is a flexible framework to encourage task-oriented dialog system research & development
Learning to Modulate pre-trained Models in RL (Decision Transformer, LoRA, Fine-tuning)
RLLaVA is a user-friendly framework for multi-modal RL research and optimized for resource-constrained teams.
Generative framework for crystal structure prediction and de novo generation of inorganic crystals.
共 233 条 · 第 10 / 12 页