llm-inference
197 个项目 · ⭐ 317.4kFineTune LLMs in few lines of code (Text2Text, Text2Speech, Speech2Text)
[arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement
The Next-Gen Database for AI—an infrastructure designed for data and AI. As the MySQL of the AI era.
Run any Large Language Model behind a unified API
AI stack for interacting with LLMs, Stable Diffusion, Whisper, xTTS and many other AI models
🦙 Free and Open Source Large Language Model (LLM) chatbot web UI and API. Self-hosted, offline capable and easy to setup.
A blazingly fast, privacy first & OPEN AI Chat Interface
⚡️ A fast and flexible PyTorch inference server that runs locally, on any cloud or AI HW.
🪶 Lightweight OpenAI drop-in replacement for Kubernetes
Large Language Model (LLM) Inference API and Chatbot
A High-Performance LLM Inference Engine with vLLM-Style Continuous Batching
共 197 条 · 第 5 / 10 页