A pure-Go, single-file AI memory and knowledge graph library and plugin.
Golang framework to build an AI that can understand and speak back to you, and everything else you want
🚢 Yet another operator for running large language models on Kubernetes with ease. Powered by Ollama! 🐫
Kubernetes Copilot powered by AI (OpenAI/Claude/Gemini/etc)
A lightweight, production-ready RAG (Retrieval Augmented Generation) library in Go.
A command-line interface (CLI) for Google Gemini
The implementation of Model Context Protocol (MCP) server for VictoriaMetrics
Open-source AI customer support system. AI-first support, human-ready operations.
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
共 1351 条 · 第 61 / 68 页