WHO LET THE AGENTS ACT · LABS FOR SECURING AI AGENTS
Find the boundary that breaks
Explore nine realistic agent-security failures, then compare how prompt instructions, model behavior, and application controls change the outcome.
01
Choose a postureSee how the same agent behaves with different defenses.
02
Send a requestUse a preset or write your own natural-language message.
03
Read the evidenceFollow the trace to see what the model and application did.
THE LAB
Nine ways an agent can cross a boundary
Choose a scenario to change the data, tools, and security boundary underneath the same comparison workflow.
LAB CONDITION · APPROVAL SERVICE
● HEALTHY
Approval requests return the database-backed policy decision.
AGENT CONSOLEGive the bank agent a request to investigate.
⌘ ENTER
AGENT RESPONSEOutcome, response, and comparison results appear here.
WAITING
→
Run or compare this request
Use one mode for its full execution trace, or compare all three security postures.
AGENTIC MECHANISM
MODEL DECIDES
SECURITY BOUNDARY
VIEW LIVE AGENT CONFIGURATION
SCENARIO DATABASE & POLICY
ACTIVE SYSTEM PROMPT
CONTROLS
EVIDENCE
RUN ARTIFACTS
Session state expires after 30 minutes of inactivity.