by openguardrails · v0.0.4
Free Open-source AI guardrails for contextually protecting against prompt attacks, content safety and sensitive data leakage.
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
Author: openguardrails
Version: 0.0.4
Type: tool
The OpenGuardrails plugin provides AI moderation and security tools for Dify applications.
It is open-source, free, context-aware, and designed for AI application developers.
OpenGuardrails focuses on:
Prompt Attack Detection
Jailbreaks, prompt injection, role-playing, rule bypass
Content Safety Detection
Context-aware detection for content safety
Sensitive Data Leakage Prevention PII, Business Secrets。
19 Risk Categories:
Open Source Repositories:
OpenGuardrails - Check Prompt
Input: prompt (user input to the model)
Output:
id:
type: string
description: "Unique identifier for the guardrails check"
overall_risk_level:
type: string
description: "Overall risk level: no_risk, low_risk, medium_risk, high_risk"
suggest_action:
type: string
description: "Suggested action: pass, reject, replace"
suggest_answer:
type: string
description: "Suggested alternative answer if action is replace, empty string if not applicable"
categories:
type: string
description: "Risk categories, separated by commas"
score:
type: number
description: "Detection probability score (0.0-1.0)"
OpenGuardrails - Check Response Contextual
Input: prompt (user input) + response (AI output)
Output:
id:
type: string
description: "Unique identifier for the guardrails check"
overall_risk_level:
type: string
description: "Overall risk level: no_risk, low_risk, medium_risk, high_risk"
suggest_action:
type: string
description: "Suggested action: pass, reject, replace"
suggest_answer:
type: string
description: "Suggested alternative answer if action is replace, empty string if not applicable"
categories:
type: string
description: "Risk categories, separated by commas"
score:
type: number
description: "Detection probability score (0.0-1.0)"
To use OpenGuardrails, you need an API Key:

Use OpenGuardrails’ OGR Check Prompt and OGR Check Response Contextual tools to protect the input and output of large language models.

For more details, workflows, and best practices, please visit:
If you encounter issues, feel free to open an Issue on GitHub.
For business cooperation, please contact thomas@OpenGuardrails.com