mlx
95 个项目 · ⭐ 185.3k🤖✨ChatMLX is a modern, open-source, high-performance chat application for MacOS based on large language models.
A modular Swift SDK for audio processing with MLX on Apple Silicon
MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. It implements OpenAI-compatible API endpoints, enabling seamless integration with existing OpenAI SDK clients while leveraging the power of local ML inference.
Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.
The easiest way to run the fastest MLX-based LLMs locally
PMetal: high-performance Apple Silicon framework for local LLM inference, LoRA/QLoRA fine-tuning, serving, quantization, and MLX/Metal acceleration.
Phi-3, -3.5, and -4 for Mac: Locally-run Vision and Language Models for Apple Silicon
Blazing fast whisper turbo for ASR (speech-to-text) tasks
Private, local transcription and system-wide dictation for Apple Silicon.
Local LLM Testing & Benchmarking for Apple Silicon
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
M-Courtyard: Local AI Model Fine-tuning Assistant for Apple Silicon. Zero-code, zero-cloud, privacy-first desktop app powered by Tauri + React + mlx-lm.
共 95 条 · 第 2 / 5 页