Biopharma Bench V0.1 results: an AI agent benchmark for pharmaceutical regulatory and CMC work, with scores on 71 tasks and 752 criteria.
2 comments
connorsun332 days ago
Really interesting Levi, what are the implications for the AI race between the US and China? And I wonder if this can lead to collaboration between companies for the purpose of advancing drug development for the betterment of humanity in general.
yujin5125 days ago
!
Read the full thread on Hacker News →
Related stories
- GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligenceartificialanalysis.aiHacker News · 47 points · about 14 hours ago
- Hacker News · 4 points · 6 days ago
- Artifical Analysis - GPT-6.1 Sol Replaces 6 Sol After 7 Daysartificialanalysis.aiHacker News · 2 points · 1 day ago
- OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.comArs Technica · 0 points · 1 day ago
- Hacker News · 2 points · 7 days ago
- Addendum to GPT-6 Astra System Card: GPT-6.1 Soldeploymentsafety.openai.comHacker News · 3 points · 1 day ago