Unify Efficient Fine-tuning of RAG Retrieval, including Embedding, ColBERT, ReRanker.
Real-time two-way speech translation for bilingual meetings — auto-detects the spoken language and translates both directions, cloud or fully offline on-device. Desktop (Windows · macOS · Linux) + browser extension (Chrome · Edge) for Zoom, Meet, Teams & any app.
DataDreamer: Prompt. Generate Synthetic Data. Train & Align Models. 🤖💤
Evaluate your LLM's response with Prometheus and GPT4 💯
A desktop multi-agent harness built with Rust, Tauri, and React, powered by langgraph-rust.
A simple, locally hosted Web Search MCP server for use with Local LLMs
An open-source, PyTorch-like runtime for dynamic multi-agent and multi-session workflows.
Your AI forgets. This remembers. Spec-driven coding harness for vibecoders, product owners, CEOs and real builders — self-improving context memory, 15 agents, 33 skills working with /goal, agent-team, & workflow on autopilot loops with 0 need for human gate. Kills context rot, ships features, not spaghetti. Claude Code & Codex. Any stack
[ICLR 2026] LightMem: Lightweight and Efficient Memory-Augmented Generation
AIConfig is a config-based framework to build generative AI applications.
Chat with your documents using local AI
Chat with AI large language models running natively in your browser. Enjoy private, server-free, seamless AI conversations.
Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM
共 3778 条 · 第 56 / 189 页