SUPERWISE® AMP → Guardrails

The AI guardrails platform that blocks failures before they reach users

Enterprise AI guardrails are runtime product controls that evaluate live AI requests and enforce policy before a response reaches a user.Framework guardrails define expectations, and code guardrails protect one application path. SUPERWISE runtime checks apply policy consistently across governed AI traffic.

What the AI guardrails platform does

Guardrails are the runtime layer of the SUPERWISE AI governance platform. Each request and response routed through the configured guardrail is checked against its configured rules for PII, toxicity, restricted topics, and jailbreak attempts. A detected violation is blocked before the user sees it.

On the platform, guardrails check PII, toxicity, topics, and jailbreak attempts on inputs and outputs. Through Sentinel, the AI gateway, the AI tools your team already uses get the secrets, PII, and toxicity guardrails on outgoing prompts.

See it work

Guardrails in action

superwise-guardrail-monitor
● Monitoring active...

Real-time policy checks

Built-in protections

Pre-configured guardrails for common risks, plus the flexibility to define your own

PII Detection

An SSN in a prompt or an agent's reply. Caught and redacted.

Toxicity Filter

Your agent just insulted a customer. Prevent harmful responses in real time.

Topic Restriction

User tried to make your agent discuss competitors. Topic blocked.

Jailbreak Detection

Attacker tried to override system prompt. Attempt logged and rejected.

Competitor Check

Agent almost named a competitor. Blocked before the word left.

Protection at every layer

1

Input Validation

Detect prompt injection, PII, and malicious inputs before model execution.

2

Runtime Evaluation

Evaluate AI outputs against your policies in real-time. Block, modify, or flag violations automatically.

From the docs

from superwise_api.models.guardrails.guardrails import ToxicityGuard

high_toxicity_guard = ToxicityGuard(name="High toxicity guard", tags=["input"], threshold=0.8, validation_method="sentence")

results = sw.guardrails.run_guardrules(tag="input", guardrules=[high_toxicity_guard], query="Heck ye")
for result in results:
    if not result.valid:
        print(f"Guardrail '{result.name}' was violated.")
        print(f"Violated message: {result.message}")

Guardrails stop threats in real-time

Watch SUPERWISE detect and block PII leaks before they ever reach the LLM. See how a patient SSN gets stopped cold.

Related capabilities

As part of SUPERWISE AMP, guardrails work together with other capabilities

AI guardrails platform questions

What buyers ask when comparing AI guardrails platforms

Deploy your first guardrail free

Deploy in minutes. Stop your next breach before it happens.