vllm
351 个项目 · ⭐ 216.0kTopicGPT: A Prompt-Based Framework for Topic Modeling [NAACL'24]
Low latency JSON generation using LLMs ⚡️
Setup and run a local LLM and Chatbot using consumer grade hardware.
Versatile Almost Local, Eventually Reasonable Assistant 🔫
Qwen3.5-122B-A10B on DGX Spark: 28.3 → 51 tok/s (+80%)
Implementation for FP8/INT8 Rollout for RL training without performence drop.
☸️ Easy, advanced inference platform for large language models on Kubernetes. 🌟 Star to support our work!
AI-powered offensive security testing using autonomous agents, directly in your terminal.
Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.
Home Assistant LLM integration for local OpenAI-compatible services (llamacpp, vllm, etc)
A CPU Realtime VLM in 500M. Surpassed Moondream2 and SmolVLM. Training from scratch with ease.
共 351 条 · 第 4 / 18 页