Create a project
Name the agent, describe its tools and choose whether policy starts in monitor or enforce mode.
About 2 minutesA fictional support agent receives a hidden instruction inside a customer email and attempts a £2,500 dry-run refund. Follow the complete evidence chain from declared controls to retest and deployment decision.
AgentRiskLayer connects the agent, policy decision, evidence and remediation in one repeatable workflow.
Name the agent, describe its tools and choose whether policy starts in monitor or enforce mode.
About 2 minutesSend prompts, outputs and proposed tool calls to AgentRiskLayer before your application acts.
One API integrationAllow safe actions, block dangerous patterns and require a server-issued approval bound to the exact sensitive operation.
Versioned policySee blocked events, assign fixes, retest the same attack and retain signed privacy-safe evidence.
Auditable resultIt checks the request against your policy. A safe action can continue. A dangerous action is blocked. A sensitive action can pause until an authorised person approves the exact customer, amount and expiry.
Community includes one project and 10,000 Guard decisions each month. No card is required.