A vulnerability scanner for LLMs: it runs a model through a set of probes for prompt injection, data leakage, toxicity and other failures. Works as a command-line tool.
Securitydata-privacy-stack
presidio
A framework that finds and masks personal data in text, images and tables. It detects the data, then redacts or replaces it so it can be safely sent to a model.
Who it is for
For developers who send user data to LLMs and need to strip personal details first.
How to start
- Open Installing Presidio in the README and pick a method: pip, Docker or from source
- Follow the Getting started guide on the docs site
- Try text anonymization on a sample of your own data
Steps are taken from the README. Check the current version in the repository before running them.
Stars over the last 30 days
Author's description
An open-source framework for detecting, redacting, masking, and anonymizing sensitive data (PII) across text, images, and structured data. Supports NLP, pattern matching, and customizable pipelines.
More in «Security»
A Python framework that checks LLM inputs and outputs with ready validators from Guardrails Hub and helps produce structured answers. It catches risks before a reply reaches the user.
NVIDIA's toolkit for programmable rails in LLM-based conversational systems. Rules live in config and control what the bot talks about and does.
Microsoft's framework for proactively finding risks in generative AI systems. It helps automate red teaming against models and applications.
Meta's generative AI safety project: the Llama Guard and Prompt Guard filter models, the Code Shield scanner and the CyberSec Eval test suites. It pairs defense with attack-based testing.
A security scanner for AI agents, MCP servers and skills. It checks configs and SKILL.md files for risky spots.
