Runnable question sets for TypeSafe's Jev, an eval harness with measured CLINC150 results, and a linter for the request shapes the API silently mis-reads. - chr-kelly/jev-cookbook
#clinc150#testing#generative#inputs#generative decision#clinc150 inputs#decision model#generative decision model
0 comments
No comments yet.
Read the full thread on Hacker News →
Related stories
- Testing the Hard Stuff and Staying Saneyoutube.comLobsters · 6 points · over 7 years ago
- Property-Based Testing Against a Model of a Web Applicationconcerningquality.comLobsters · 7 points · over 3 years ago
- Lobsters · 2 points · over 7 years ago
- Testing sudo-rs and improving sudo along the wayferrous-systems.comLobsters · 7 points · about 3 years ago
- Lobsters · 53 points · 8 months ago
- More tools for testing SQL dialectsbuttondown.comLobsters · 2 points · 6 months ago