Tag
#guardrails
5 posts tagged guardrails.
- Defense
LLM Jailbreak Defense Techniques: A Practitioner's Guide
A technical breakdown of LLM jailbreak defense techniques, from input filtering and model-level hardening to multi-agent architectures and OWASP mitigations.
- AI Security
Prompt Injection Detection Tools: The 2026 Landscape for LLM Defenders
Prompt injection tops the OWASP LLM Top 10, and the tool market has consolidated fast. A field guide to Llama Prompt Guard 2, Azure Prompt Shields, Lakera Guard, NeMo Guardrails, and the open-source options that survived.
- news
How LLM Chatbots Leak Data Through Their Own Rendered Output
A recurring AI-security finding: an injected instruction makes the model emit a markdown image whose URL carries the user's data to an attacker server.
- news
Indirect Prompt Injection: The Agent Era's Default Vulnerability
As LLM agents gained tools and memory, the dangerous injection stopped coming from the user and started coming from the data the agent reads.
- news
The OWASP LLM Top 10 (2025) Changed More Than the Numbering
The 2025 revision of the OWASP Top 10 for LLM Applications added system-prompt leakage and vector/embedding weaknesses, and reframed the supply-chain