Llm-Firewall

Latest
Model-as-a-Judge Has Its Own Prompt Injection Problem
Research shows attackers can manipulate AI judges. Smaller models, larger models, and typed decisions each carry security and performance tradeoffs.

The First AI-Run Intrusion Was a Visibility Failure
A frontier-model evaluation produced an autonomous intrusion. Milgram's replay shows …

Prompt Injection Is the SQL Injection of the AI Era
Prompt injection now tops the OWASP LLM risks and can't be patched away like SQLi. …

Agentic AI Just Multiplied Your Attack Surface
Agentic AI turns every model call into an action, expanding the attack surface. …

Turning Data-Leak Risk Into a Metric Your Board Understands
Reframes AI data-leak exposure as a board-level metric, using 2025 breach data to argue …

The Cost of Context: Why Bigger Prompts Aren’t Better
Longer prompts cost more and answer worse; the fix is treating the AI traffic boundary as …