RAG app for patent similarity search with chatgpt llm over google patents
Pull high-quality, efficient embeddings for PubMed, arXiv and Wikipedia from Huggingface and use for local LLM inference/Retrieval Augmented Generation (RAG)
Batch Deployment for Document Parsing with AWS Batch & Qwen-2.5-VL
Bleeding edge vLLM Docker image for the NVIDIA DGX Spark (GB10 / sm_121a).
The purpose of the "Meta Agent with More Agents" project is to dynamically solve complex queries by breaking them down into smaller tasks and assigning each to specialized AI agents. The Meta Agent coordinates the process, leveraging a ReAct Agent for tool-based tasks and a Chain of Thought Agent for reasoning-based tasks. The system's flexibility.
A SillyTavern extension providing a multi-tier memory and narrative context system for AI roleplay. Tracks character facts, relationship history, per-character knowledge and secrets, entity state, scene history, story arcs, and rolling summaries - extracted automatically so your AI stays coherent no matter how long the story runs.
Code repository for AI Builders Bootcamp #2
[ICLR'24 Spotlight] DP-OPT: Make Large Language Model Your Privacy-Preserving Prompt Engineer
A user-customized bot for your slack channels using LLMs, Tools and Documents
Implementation of 12 AI agents evaluation techniques
This repository contains a web application designed to execute relatively compact, locally-operated Large Language Models (LLMs).
Interpretable text embeddings by asking LLMs yes/no questions (NeurIPS 2024)
共 3770 条 · 第 153 / 189 页