amd
29 个项目 · ⭐ 229.3kA high-throughput and memory-efficient inference and serving engine for LLMs
A bundler for javascript and friends. Packs many modules into a few bundled assets. Code Splitting allows for loading parts of the application on demand. Through "loaders", modules can be CommonJs, AMD, ES6 modules, CSS, Images, JSON, Coffeescript, LESS, ... and your custom stuff.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
GPU & Accelerator process monitoring for AMD, Apple, Huawei, Intel, NVIDIA and Qualcomm
Create graphs from your CommonJS, AMD or ES6 module dependencies
Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.
Fast and lightweight x86/x86-64 disassembler and code generation library
QualityScaler - image/video AI upscaler app
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
web UI for GPU-accelerated ONNX pipelines like Stable Diffusion, even on Windows and AMD
Stable Diffusion Docker image preconfigured for usage with AMD Radeon cards
共 29 条 · 第 1 / 2 页