I ran 5 prompt-injection attacks against my chatbot's system prompt. Here's what broke.

Everyone shipping an LLM feature worries about jailbreaks, but 'is our system prompt actually resilient?' gets answered by vibes. So I built a tester for it.

Read Original

Related