Show HN: Jevstiller – Distill Jev into a local model, with a disagreement bound
tgluck
62 points
14 comments
September 29, 2026
Related Discussions
Found 5 related stories in 89.7ms across 8,041 title embeddings via pgvector HNSW
- Show HN: JevBench, a reproducible benchmark for typed decision models florianstandhar · 88 pts · September 22, 2026 · 61% similar
- Show HN: Distill and serve models with frontier quality for half the cost SilenN · 42 pts · July 26, 2026 · 59% similar
- Show HN: jevals – replacing LLM judges with typed Jev decisions gbayomi · 15 pts · September 20, 2026 · 58% similar
- Jeeves. Reasoning improves Jev-like decision models nicowaltz · 234 pts · September 29, 2026 · 54% similar
- Introducing System One Models and Jev albelfio · 1080 pts · September 15, 2026 · 54% similar
Discussion Highlights (3 comments)
tgluck
Author here. This puts a proxy in front of repeated Jev classification calls. At first everything goes to Jev; from Jev's answers it trains a small head on frozen sentence embeddings, picks a confidence threshold with an exact finite-sample bound so that at most 2% of all requests get an answer Jev wouldn't have given, and then answers the confident share locally at ~15 ms on a CPU. A permanent 2% audit keeps checking; if agreement breaks, everything falls back to Jev and it retrains. Known limits: agreement is not accuracy (if Jev is wrong, so is the local model); coverage tracks how consistent Jev itself is (22% on noisy tweet tasks, 80% on news); it speaks Jev's API only, an OpenAI-compatible front is on the roadmap. Since 0.4.0 the guarantee can also cover "would Jev have been unsure", which matters if your code routes low-confidence answers to review. Apache 2.0.
gingersnap
Is the local model similar to model2vec?
ricardobeat
This will only work for simple text classification tasks, which is the least interesting possible use of Jev.