Agent♥︎Age
Catalog

io.github.Vishisht16/humane-proxy

Official

by Vishisht16 · Python

AI safety middleware — detects self-harm and criminal intent in LLM prompts.

io.github.Vishisht16/humane-proxy (MCP) Server

AI safety middleware that detects self-harm ideation and criminal intent in LLM prompts. HumaneProxy sits between users and any LLM, intercepting relevant messages before the LLM ever sees them, and alerts you through preferred channels while responding with care.

🛠️ Key Features

  • Detects self-harm ideation in LLM prompts
  • Detects criminal intent in LLM prompts
  • Intercepts messages before they reach an LLM
  • Alerts via preferred channels
  • Provides a care-oriented response

🚀 Use Cases

  • Adding guardrails to existing LLM integrations
  • Monitoring prompts for self-harm prevention signals
  • Monitoring prompts for crime-prevention signals

⚡ Developer Benefits

  • Lightweight, plug-and-play middleware for Python-based setups
  • Works as MCP middleware (“mcp” topic) in an LLM request flow
  • Centralized interception and handling of sensitive prompt content

⚠️ Limitations

  • Description only specifies detection and interception; tool interfaces and counts are not provided in the available data.
safetycontent filteringharm detectionllm securityprompt screeningrisk mitigation

Topics

ai-safetycrime-preventionguardrailsllmmcpmiddlewarepythonself-harm-prevention

Related servers

More in Security