ggml
30 个项目 · ⭐ 25.4kDiffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
[Unmaintained, see README] An ecosystem of Rust libraries for working with large language models
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model
Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization
This custom_node for ComfyUI adds one-click "Virtual VRAM" for any UNet and CLIP loader as well MultiGPU integration in WanVideoWrapper, managing the offload/Block Swap of layers to DRAM *or* VRAM to maximize the latent space of your card. Also includes nodes for directly loading entire components (UNet, CLIP, VAE) onto the device you choose
Port of MiniGPT4 in C++ (4bit, 5bit, 6bit, 8bit, 16bit CPU inference with GGML)
CLIP inference in plain C/C++ with no extra dependencies
Large Language Models for All, 🦙 Cult and More, Stay in touch !
llama.cpp (GGUF LLMs) and llava.cpp (GGUF VLMs) for ROS 2
Booster - open accelerator for LLM models. Better inference and debugging for AI hackers
🪶 Lightweight OpenAI drop-in replacement for Kubernetes
共 30 条 · 第 1 / 2 页