Papers and resources related to the security and privacy of LLMs 🤖
-
Updated
Jun 8, 2025 - Python
Papers and resources related to the security and privacy of LLMs 🤖
[NeurIPS D&B '25] The one-stop repository for LLM unlearning
The fastest Trust Layer for AI Agents
An Execution Isolation Architecture for LLM-Based Agentic Systems
Make Zettelkasten-style note-taking the foundation of interactions with Large Language Models (LLMs).
Semantic PII Masking & Anonymization for LLMs (RAG). GDPR-compliant, reversible, and context-aware. Supports LangChain & OpenAI
🔒 Detect security leaks in AI-assisted codebases. Static analysis tool for Python & JS/TS with cross-file taint tracking.
Example of running last_layer with FastAPI on vercel
Local-first safety/control harness for AI coding agents with reviewable handoff packets, route audits, and bounded cloud escalation.
Local proxy for Claude Code that redacts sensitive data (names, emails, IDs, keys) before it leaves your machine and restores it in the reply
🛡️ Scrimward — local, fail-closed redaction proxy for AI coding tools (Claude Code · Codex · Cursor · Copilot · Gemini …): mask secrets, PII & images on your machine before they reach the cloud. (Early development.)
Add a description, image, and links to the llm-privacy topic page so that developers can more easily learn about it.
To associate your repository with the llm-privacy topic, visit your repo's landing page and select "manage topics."