We didn't make the models smarter. We built the thing that catches them confidently wrong — and it caught us too.
DEV Community·3 points·bryanw·10 days ago·dev.to
The one-line version We ran five current frontier models over a set of documented-failure...
Read the full article at dev.to →
Related stories
- My prompt-injection fix caught 0 of 20 attacks. The part I almost didn't build caught all of them.dev.toDEV Community · 3 points · 4 days ago
- Hacker News · 40 points · about 14 hours ago
- Hacker News · 1 points · 8 days ago
- Hacker News · 2 points · 9 days ago
- Hacker News · 2 points · 8 days ago
- Belay – See what keeps going wrong in your Claude Code and Codex sessionsgetbelay.vercel.appHacker News · 1 points · 9 days ago