mixture-of-experts
26 个项目 · ⭐ 62.1kDeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Decentralized deep learning in PyTorch. Built to train models on thousands of volunteers across the world.
Run Mixtral-8x7B models in Colab or consumer desktops
Codebase for Aria - an Open Multimodal Native MoE
A library for easily merging multiple LLM experts, and efficiently train the merged LLM.
Implementation of Soft MoE, proposed by Brain's Vision team, in Pytorch
PyTorch library for cost-effective, fast and easy serving of MoE models.
🚀 LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training
Nodes to run Hunyuan Image 3 locally with BF16 and NF4 quantized options in Comfyui
hummingbird is a lightweight, zero dependency runtime for massive open source Mixture of Experts (MoE) language models. It unifies SSD, RAM, and VRAM into a single intelligent memory hierarchy, enabling inference of models like GPT-OSS 120B, GLM, DeepSeek, Qwen, and more on consumer hardware
共 26 条 · 第 1 / 2 页