llama-cpp
104 个项目 · ⭐ 108.6k1-Click LLM Server on Your Phone — no Termux needed! 无需Termux,一键让你的手机变成LLM服务器!
VindexLLM is a pure Delphi, GPU-powered LLM inference engine that uses Vulkan compute shaders to run GGUF models entirely on the GPU. It performs full transformer inference without relying on Python, CUDA, or other external runtimes, requiring only vulkan-1.dll, which is typically included with modern GPU drivers.
Android native AI inference library, bringing text, image, video, STT, TTS inference
😈 ImpAI is an advanced role play app using large language and diffusion models.
A C++ implementation of Open Interpreter. / Open Interpreter 的 C++ 实现
◉ Universal Intelligence: AI made simple.
Windows 本地 AI 语音输入:Qwen3-ASR CPU/GPU 双运行形态、本地屏幕 OCR 上下文、热词纠错与可选 AI 润色。
A terminal coding agent, and a Python SDK for embedding on-device models in your own apps.
Android keyboard with local AI (Ollama, Whisper, MCP) or cloud (Gemini, Groq, OpenAI)
One GPU. Full LLM workflow. Real benchmarks. No cloud required.
共 104 条 · 第 4 / 6 页