This is the sixth article in a series about Agent Experience (AX): the practice of making AI coding agents work correctly with your technology. The series
#coding#openai#getting better#openai models#new openai#models getting#openai models getting#new openai models
3 comments
Zeruxe1 day ago
What even makes it worse is that they always make the models worse a few days after release, this is a wide problem across all the companies
bmoathn1 day ago
do you know of anyone that's put out quantified studies on this? I'd be super curious to see. Like benchmarks at release vs 1wk vs 1mo?
Zeruxe1 day ago
I would 100% be down to research it, I just know it but I cant prove it because I can feel I have agents running 24/7 almost and almost every time at the beginning of the launch they are great and they gradually decrease everyday...
Read the full thread on Hacker News →
Related stories
- Hacker News · 55 points · 4 days ago
- The Verge · 0 points · 1 day ago
- OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.comArs Technica · 0 points · 1 day ago
- The Verge · 0 points · 8 days ago
- Ars Technica · 0 points · 2 days ago
- Ars Technica · 0 points · 6 days ago