Agent♥︎Age
Catalog

io.github.IgorGanapolsky/rlhf-feedback-loop

Official

by IgorGanapolsky · JavaScript

RLHF feedback loop for AI agents. Capture signals, promote memories, block mistakes, export DPO.

io.github.IgorGanapolsky/rlhf-feedback-loop (MCP Server)

This MCP server provides an “RLHF feedback loop for AI agents.” It is associated with ThumbGate, described as a “self-improving pre-action firewall for AI coding agents,” intended to use thumbs up/down signals to manage agent actions before execution. The project emphasizes pre-action checks and blocking mistakes, with capabilities to export DPO.

🛠️ Key Features

  • RLHF feedback loop for AI agents
  • ThumbGate pre-action firewall (👍/👎 signaling)
  • Capture signals and “promote memories”
  • Block mistakes and run pre-action checks
  • Export DPO

🚀 Use Cases

  • Improving AI coding agent reliability
  • Preventing directory wipes, key leaks, or broken code from wrong tool calls
  • Applying guardrails to agent tool execution
  • Agent feedback and self-improvement loops

⚡ Developer Benefits

  • Agent guardrails via thumb-based infrastructure firewall
  • Pre-action checks to reduce harmful or incorrect tool calls
  • Integration with MCP server workflows
  • Support for DPO export

⚠️ Limitations

  • The provided description and excerpt explain behavior at a high level; specific implementation details are not included in the available source.

Topics

ai-safetydeveloper-toolsai-agentsguardrailsthompson-samplingagent-reliabilityclaude-codecodexcursorfeedback-loopgeminimcpmcp-serveropencodepre-action-checksthumbgateself-improving-agentsamp