trustmebro
Bypass LLM guardrails with fabricated tool output
Visit trustmebro ↗
link: rel="ugc noopener"
trustmebro is a tool designed to test LLM security by simulating tool outputs to confuse language model guardrails. It demonstrates how models can be manipulated through fake function call responses.
Available on GitHub, this project is intended for security research and testing purposes, allowing developers to understand potential vulnerabilities in LLM safety mechanisms.
Discussion