4 comments
mbeavitt2 days ago
I'm not going to write some big blog post here, but this chat log is quite entertaining. I wrote a fundamental string algorithm in C and GPT 5.6 Sol on medium effort hallucinated not one but two correctness issues with the code!
Maybe this is not surprising to others but it's been a long while since I caught one of these models making such a glaring error.
chris_money2022 days ago
I'm pretty sure there is a completely different model for Chat vs Coding.
Like if you deploy 5.6 on azure foundry there is a chat model and a coding/reasoning model
mbeavitt2 days ago
I learned something new today
Read the full thread on Hacker News →
Related stories
- Lobsters · 86 points · about 1 year ago
- GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligenceartificialanalysis.aiHacker News · 47 points · about 15 hours ago
- Hacker News · 4 points · 6 days ago
- Artifical Analysis - GPT-6.1 Sol Replaces 6 Sol After 7 Daysartificialanalysis.aiHacker News · 2 points · 1 day ago
- Hacker News · 14 points · 8 days ago
- OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.comArs Technica · 0 points · 1 day ago