Frontier language models are rarely used in clinical workflows because the realistic, longitudinal benchmarks needed to develop them are scarce. Real electronic health record (EHR) data cannot be openly shared due to…
0 comments
No comments yet.
Related stories
- Hacker News · 1 points · 3 days ago
- DEV Community · 7 points · 3 days ago
- Hacker News · 2 points · about 12 hours ago
- Lobsters · 76 points · over 1 year ago
- Hacker News · 1 points · 10 days ago
- JevBench: Benchmark for Jev-Class Modelsbenchmarkheaven.comHacker News · 2 points · 10 days ago