DavidCarliez/trustmebro: Bypass llm guardrails by confusing it with fabricated tool output
github.com | tool | #ai-security | #open-source | #red-team | #llm-security | #ai-agents | #guardrail-bypass | #go | #llm | #guardrails | #tool-calling | #path-shim
Summary
trustmebro bypasses LLM guardrails by spoofing tool output: PATH shims intercept commands like dig and return rewritten stdout, deceiving agents that trust shell tools, with a JSONL audit log.
- Collected
Skip to content