gpu
164 个项目 · ⭐ 715.5k🪶 Lightweight OpenAI drop-in replacement for Kubernetes
One-command vLLM installation for NVIDIA DGX Spark with Blackwell GB10 GPUs (sm_121 architecture)
Real-time hardware and LLM inference monitoring — GPU, CPU, memory, and vLLM metrics streamed to a dashboard.
Self-host a ChatGPT-style web interface for Ollama 🦙
Extensible generative AI platform on Kubernetes with OpenAI-compatible APIs.
Neural networks library for machine learning on PHP
AMD ROCm Installation Guide on RX 6600 XT + TensorFlow and PyTorch
🚀 100% local RAG system with one-command setup. Your data never leaves your server. Apache-2.0
:dart: Gradient Accumulation for TensorFlow 2
A high-performance RDMA distributed file system for fast LLM Inference and GPU Training.
First open-source TurboQuant KV cache compression for LLM inference. Drop-in for HuggingFace. pip install turboquant.
共 164 条 · 第 6 / 9 页