2 comments
Reubend9 days ago
It's a Qwen fine tune that's benchmaxxing a few sets of question types, and it's obviously nowhere near beating Opus 3.8 in real usage. Nothing to see here.
Normally, I'd give them a pass because at least it's open source, but then I read
> Raising pre-seed. So we welcome all angels and vcs
Which basically gives away that they're trying to fool investors into giving them money by exaggerating their results.
kadoban9 days ago
There is, IMO a lot of value in fine tuning existing models. If that gets them funded, great, who cares, right?
This is not their first release. BTL-2 (or -3 maybe? I forget) looked kind of interesting but it was annoying to run, needed a patch on llama.cpp, looks like they're improving their tooling.
I'm certainly interested in how this actually performs.
Read the full thread on Hacker News →
Related stories
- The Verge · 0 points · 9 days ago
- Hacker News · 2 points · 8 days ago
- The Economics of Open-Weight Inferencedata.ornn.comHacker News · 36 points · 9 days ago
- Hacker News · 1 points · 8 days ago
- Beats 360 Over-Ear Headphone from Beatsbeatsbydre.comHacker News · 1 points · 8 days ago
- DEV Community · 22 points · 11 days ago