← 返回专题广场
rl
31 个项目 · ⭐ 73.8k21
2025-08-07
最近推送
22
Python library for solving reinforcement learning (RL) problems using generative models (e.g. Diffusion Models).
Python
⭐ 222
⑂ 13
Apache-2.0
· 2025-02-18推送
2025-02-18
最近推送
23
2 天前
最近推送
24
2025-06-28
最近推送
25
RLLaVA is a user-friendly framework for multi-modal RL research and optimized for resource-constrained teams.
Python
⭐ 58
⑂ 6
Apache-2.0
· 28 天前推送
28 天前
最近推送
26
[AAAI 2026] D²PPO: Diffusion Policy Policy Optimization with Dispersive Loss.
Python
⭐ 42
⑂ 4
MIT
· 2025-11-22推送
2025-11-22
最近推送
27
2025-10-21
最近推送
28
Awesome-HCI (Ubiquitous, LLM, MLLM, Agent, RAG, Embodied-AI, RLHF)
Python
⭐ 25
⑂ 2
· 2026-03-08推送
2026-03-08
最近推送
29
Context & Guide For Reinforcement Learning with Verifiable Rewards with Large Language Models
Jupyter Notebook
⭐ 21
⑂ 2
· 2025-11-04推送
2025-11-04
最近推送
30
KONASH: Train knowledge agents that search, retrieve, and reason. Based on KARL (Databricks, 2026).
Python
⭐ 17
⑂ 0
· 2026-03-23推送
2026-03-23
最近推送
31
29 天前
最近推送
共 31 条 · 第 2 / 2 页