91.5% recall at sub-millisecond latency from regex alone, 98.1% with a ML layer added — and the false positives that come with it. Full methodology, mapped to the OWASP Top 10 for LLM Applications.
1 comment
textcortex1 day ago
Can you also benchmark this one too? We recently trained it and it works well in production: https://jays.fyi/blog/open-prompt-injection-scanner-for-ai-a...
Read the full thread on Hacker News →
Related stories
- DEV Community · 27 points · 4 days ago
- Prompt Injection in the Wildcybershujin.github.ioHacker News · 2 points · 5 days ago
- Is indirect prompt injection still a big threat as models get more advanced?realarcherl.github.ioHacker News · 3 points · 6 days ago
- Meta Muse: Almost no prompt injection resistanceneuromatch.socialHacker News · 1 points · 6 days ago
- Hacker News · 6 points · 7 days ago
- Hacker News · 1 points · 2 days ago