llm-security
47 个项目 · ⭐ 171.1kOpen-source AI penetration testing tool to find and fix your app’s vulnerabilities.
NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems.
🐢 Open-Source Evaluation & Testing library for LLM Agents
[CCS'24] A dataset consists of 15,140 ChatGPT prompts from Reddit, Discord, websites, and open-source datasets (including 1,405 jailbreak prompts).
safe execution paths for agents - zero trust, zero setup, zero latency.
A secure low code deception runtime framework, leveraging AI for System Virtualization.
A powerful tool for automated LLM fuzzing. It is designed to help developers and security researchers identify and mitigate potential jailbreaks in their LLM APIs.
ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.
A security scanner for your LLM agentic workflows
AI-first security scanner. NEW in v2026.7: Claude Code compromise detection — vet .claude/ hooks, permissions & skills before you clone — plus an always-on AI attack-signature scanner and native Rust & PHP rules. Also: medusa scan --git to vet any repo, medusa secrets scan for leaked API keys. 40,000+ patterns, zero setup.
Open-source adversary emulation for AI agents and MCP servers.
共 47 条 · 第 1 / 3 页