Library · topic

Security

Prompt injection, memory poisoning, and the agent-specific attack surface that appears the moment a model can act. The through-line: assume some attacks land, and design so a fooled model still can't do damage — least privilege, broken trifectas, and approval gates.

Ask a question about security