Jeff – Jev-compatible 0.8B decision models
For the folks asking what is the point of a Jev-class model, which is supposed to be fast and cheap. Losing 10 points on MMLU along the way doesn't help. Also ios 27. this is the way, a hybrid approach where some of the pipeline will be jev like and some traditional LLM depending on the nature of the work. What about benchmark against smaller or bigger models? 9B looks too small for llm-level decisions.
Interesting bench list, what about benchmark against smaller or bigger models? 9B looks too small for llm-level decisions. Cool engineering, but 17s p90 latency kind of defeats the point of Sonnet 5.5 when Opus 5.5 is cheaper than Sonnet 5. Now I know the reason.