intel/ipex-llm
Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discrete GPU such as Arc, Flex and Max); seamlessly integrate with llama.cpp, Ollama, HuggingFace, LangChain, LlamaIndex, vLLM, DeepSpeed, Axolotl, etc.
⭐ 8.9k
⑂ 1.4k
Python
Apache-2.0
· 2026-01-29推送
8.9k
Watchers
0
贡献者
0
Commits
0
Releases
1.5k
Open Issues
2026-01-29
最近推送
原文
中文
暂无 README