reinforcement-learning
233 个项目 · ⭐ 1118.6k🏆Winning Project | ModelGate is a contract-aware AI control plane that ingests customer contracts, extracts SLA/privacy/routing constraints, and generates an OpenAI-compatible endpoint that automatically routes every request to the optimal model. Simple queries go to cheap models. Complex queries go to premium ones.
Implementation of Multi-Game Decision Transformers in PyTorch
Implementation of GATO style Generalist Multimodal model capable of image, text, RL and Robotics tasks
A Continual Learning Framework for Production LLM Agents
Machine learning robotics engineer preparation material.
Panacea is a framework for building collaborative, intelligent multi agent AI systems. The framework provides a robust infrastructure for creating and managing multiple AI agents, and enables developers and organizations to build, deploy, and optimize AI agents that work well in dynamic, complex environments.
A lightweight post-training framework for LLMs and VLMs. 51 algorithms, 38 verified models. Scales with DeepSpeed, vLLM, and Ray.
A multi-purpose repository with Sentiment Analysis of Stocks news, and Reinforcement learning based bot to trade stocks.
Multi-node distributed LLM training framework
KONASH: Train knowledge agents that search, retrieve, and reason. Based on KARL (Databricks, 2026).
共 233 条 · 第 11 / 12 页