Skip to content

Self harm guardrail - #3483

Open
jareallen213 wants to merge 5 commits into
devfrom
self-harm-guardrail
Open

Self harm guardrail#3483
jareallen213 wants to merge 5 commits into
devfrom
self-harm-guardrail

Conversation

@jareallen213

@jareallen213 jareallen213 commented Jul 22, 2026

Copy link
Copy Markdown

Description

Add form for new self-harm Guardrail (SEMOSS/Semoss#2796)

Adds form configurations for two new embedded guardrail types in the guardrail import flow:

On Topic — rejects prompts whose similarity score against a configured vector database falls below a threshold, with fields for selecting the vector engine, threshold, and nearest-neighbour limit.
Aggressive / Self-Harm — routes prompt evaluation through a configured LLM to detect aggressive or self-harm content, with fields for selecting the model engine and score threshold.

Changes Made

  • adds self-harm/aggression guardrail type to guardrail form

How to Test

  1. Create new Guardrail
  2. Check configured fields + smss file match

Notes

@jareallen213
jareallen213 requested a review from a team as a code owner July 22, 2026 20:54
@snyk-io

snyk-io Bot commented Jul 22, 2026

Copy link
Copy Markdown

Snyk checks have passed. No issues have been found so far.

Status Scan Engine Critical High Medium Low Total (0)
Open Source Security 0 0 0 0 0 issues

💻 Catch issues earlier using the plugins for VS Code, JetBrains IDEs, Visual Studio, and Eclipse.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants