Create Guardrail Policies
Overview
The Create Guardrail Policies page allows you to define enforcement rules that evaluate agent requests and responses against security, safety, and compliance criteria.
Each guardrail policy combines metadata, content filters, detection logic, and deployment scope into a single configuration that applies consistently across selected agents and servers.
Access Guardrail Policies
You can access guardrail policy configuration from the Akto console.
Navigate to the Agentic Security product.
Select Agentic Guardrails → Guardrail Policies.
The guardrail policies list displays existing policies and provides access to policy creation.
Create a Guardrail Policy
Access the Create Guardrail Form
Locate the Create Guardrail button in the top-right corner of the Guardrail Policies page.
Select Create Guardrail to open the guardrail configuration form.

Fill the Configuration Form
The configuration form is organised into multiple sections, each targeting a specific enforcement layer.
Guardrail Deployment Scope Behaviour
By default, a newly created guardrail policy applies to all agentic assets — all Agents, MCP Servers, and LLMs. To restrict enforcement, select Select Agentic Assets in the Scope section and choose the specific assets to target.
Save the Guardrail Policy
After completing the required and optional configurations:
Click on Create Policy to save the policy and applies enforcement to the selected scope.
Test guardrail behaviour in the playground
The playground allows your team to validate guardrail behaviour before updating the policy.
Enter a prompt in the Test your guardrail policy field to simulate a request against the configured guardrail policy.
The playground evaluates the prompt using the selected guardrail configuration and displays the enforcement result.
You can also use the Quick Test Prompts provided in the playground to test common scenarios such as sensitive data exposure, prompt injection attempts, or abusive language.

Playground probing helps your security team verify that guardrail conditions correctly detect violations and return the expected blocked response message before the policy is finalized.
Preview Change Impact Before Saving
When you edit an existing guardrail policy, the right-hand panel shows Change impact analysis instead of the plain Impact analysis view shown while creating a new policy. It replays recent activity through both the currently saved policy and your unsaved changes, so you can see how an edit would affect detections before you save it.
Switch between the Violations and Traffic tabs before running:
Violations replays the policy's last few recorded violations.
Traffic replays the latest agent traces.
Select Run to open the Change impact analysis window and replay that activity against both versions of the policy.
The Saved policy and Your changes counts show the total detections under each version, with a badge (e.g.
-1) highlighting the net change.The results table lists each replayed Prompt alongside two columns, Saved and Draft, each marked Detected or Missed so you can see exactly which prompts change outcome under your edits.
Use Page Size and the pagination controls to page through results, or select Export CSV to download the full comparison.
Select Close to return to the policy editor.
What’s Next
You can modify, disable, or delete existing guardrail policies after creation.
To continue, learn how to manage guardrail policies from the Manage Guardrail Policies.
Last updated













