LLM-as-a-judge uses one model to score another model's output. See how judges prompt and score, the bias risks, and when they beat human review.
0 comments
No comments yet.
Related stories
- Hacker News · 1 points · 2 days ago
- Hacker News · 1 points · 9 days ago
- An LLM Beat NetHackkenforthewin.github.ioHacker News · 2 points · 7 days ago
- Hacker News · 3 points · 9 days ago
- Hacker News · 1 points · 10 days ago
- An LLM Beat NetHackkenforthewin.github.ioHacker News · 1 points · 9 days ago