September 2026. Every number here is from the benchmarks, and bash experiments/bench.sh --no-record reruns them without an API key.
15 comments
they distill different things. model2vec distills a sentence transformer into static embeddings, so the output is a faster general-purpose encoder.
Jevstiller keeps the encoder frozen (bge-small by default) and distills Jev's decisions on one specific question into a small head on top of it
Known limits: agreement is not accuracy (if Jev is wrong, so is the local model); coverage tracks how consistent Jev itself is (22% on noisy tweet tasks, 80% on news); it speaks Jev's API only, an OpenAI-compatible front is on the roadmap. Since 0.4.0 the guarantee can also cover "would Jev have been unsure", which matters if your code routes low-confidence answers to review. Apache 2.0.
What type of head is that? What type of model is that head part of?
The head is a multinomial logistic regression: one linear layer plus softmax on top of a frozen sentence-embedding model (bge-small by default, swappable). That head is the entire local model, the encoder is off the shelf and never changes.
Drop-in yes: point TYPESAFE_BASE_URL at it and nothing else changes.
Read the full thread on Hacker News →
Related stories
- Hacker News · 1 points · 9 days ago
- Show HN: A local alternative to Jev – 94% on Banking77gist.github.comHacker News · 2 points · 4 days ago
- Hacker News · 3 points · 3 days ago
- Hacker News · 1 points · 5 days ago
- Hacker News · 7 points · 11 days ago
- Hacker News · 2 points · 7 days ago