Build your own decision model

nishtahir.com

406 points by softwaredoug 17 hours ago


sn0n - 14 hours ago

I saw Jev and without understanding it said to myself, “I can build that!” And brainstormed some weird Alex trebek jeopardy generator I called trebek, a bun ran typescript app that you give it an input, it decides if it’s a category, question or answer then generates what’s missing. Trained a SQLite-vec database on the English language for a few days with a small qwen embedding model to add vec embeds to the db then ripped the cord on the embeddings before I started feeding my llm the proper specs for Jev and now i have a cool jeopardy generator that now doubles as a Jev clone classifier with a custom v1 endpoint for system 0 or whatever it is. Level understanding here, fun experiment though ^>^ and produces useful outputs

howunfortunate - 16 hours ago

As an MLE who has been failing to get anyone interested in classifiers for many years, the hype around Jev makes me scream internally.

Yes, I get that a zero-shot classifier is more convenient than the traditional kind, it's very cool. Kind of. But then again plain LLMs have been perfectly cheap and serviceable as zero-shot classifiers for quite some time now, so again I'm back to my internal screaming.

sachaa - 11 hours ago

This is a great way to start but the perf of LLM models like Qwen are not ideal for local execution. I have adapted Laya (pure decision model) to run in a browser and I am able to get responses under 200ms. Give it a try: https://wexare-ai.github.io/browser-laya/

bob1029 - an hour ago

If you don't really care about knowing the specific probability (I think this is actually a huge info hazard), you can achieve this right now with a trivial tool calling arrangement. The benefit still stands. You are reducing the size of the action space from something that might be Turing complete (shell, code) into a multiple choice question.

softwaredoug - an hour ago

To me the Jev moment isn't about Jev / Typesafe. It's articles like these. It's the letting a million Jevs bloom.

The bubble will come from a realization that a lot of people can train these models. AI Researchers will become more diffuse, work at more companies, and building a Jev or fine-tuning an LLM isn't some trillion dollar frontier lab exercise, but increasingly just something developers do.

mastazi - 10 hours ago

So "decision model" is the new cool term for classifiers? You know, that thing that already existed decades ago

nico - 15 hours ago

This is very cool. If you are looking for something similar but more lightweight, that you can run (and train) on CPU, try out Jeffy: https://jeffyclassify.com/

On GitHub: https://github.com/nicobrenner/jeffy

utbabya - 4 hours ago

Saw the question "Where would you most likely find a bat?", it occurs to me there's an innate tension, do we want the llm to be factually correct or do we want it to be more average human like? As an average human being not a sme on the subject my first instinct answer would be cave as well. I think it's reasonable to expect trainning on the aggregate of the internet means it would arrive at the same answer.

Edit: The context of the question does indeed make it sound more like the animal bat. The other answers sound more like gotchas to me.

sva_ - 15 hours ago

Does someone have examples of interesting stuff that has been built utilizing Jev/decision models? The way this is hyped up surely there must be some good stuff?

gw0 - 4 hours ago

I had the exact same thought "I can build that!" one month ago. So I built JobFit, a CV/job-post fit-scoring typed-decision model, training pipeline, and web app that runs in your browser: https://github.com/gw0/jobfit-model

demibabs - 16 hours ago

Is simply changing the temperature so that the model appears calibrated over a particular benchmark after the fact “allowed”? Feels p-hacking esque.

nl - 5 hours ago

What is the intuition behind the temperature fitting for calibration step?

Flattening the output probability distribution curve makes some sense, but playing with the graph doesn't seem to show the 3.797 figure as the best.

And how do you curve fit for this single example?

kekebo - 10 hours ago

Just for the author: On any of my iOS 26 browsers (orion, brave, safari), a page reload occurs when the model download completes, which resets the state, so I never get to interact with the model.

onatm - 6 hours ago

Nice read! I also gave a shot building one on Gemma3 and Gemma4. It was a fun exercise. I’m sharing it if anyone interested, Gemma4 based one is on a branch: https://github.com/onatm/gev

fuddle - 12 hours ago

I'm looking forward to a "Build a Decision Model (From Scratch)" book.

mrkn1 - 6 hours ago

Great guide! It works, made a 500MB model for CPU based off qwen 3.5 0.8B. It plays maze games like Pac man https://github.com/kouhxp/gutsy

ursaguild - 16 hours ago

This is really cool to see. Being able to play the token generation was awesome. Amazing job with breaking down how to think about these models. This made the idea of Jev/decision models really easy to grasp for me. The idea of calibrating the model was helpful. I thought this was a great overview.

ford - 14 hours ago

These are neat - and the source of the many Jev clones we've seen. I think their recent funding round is in part because of their algorithms/data. It remains to be seen if that's a big enough edge to be worth 1 billion+ dollars

vezycash - 9 hours ago

A lot of LLM "thinking" involve choosing between alternatives.

Offload the decision-making parts of LLM reasoning to a jev like model.

uzerkayat - 8 hours ago

Hey I have one question Is there any way we can also make a model, that can take a first decision about voice policing let's say I have speech to text tool like a whisper flow and if we want to make a voice policing decision model that can take first decision so that the voice policing works fast do it that can work out can you guys answer is.

2bitencryption - 9 hours ago

this is just constrained generation? I thought the latest crop of decision models (inspired by Jev) do something fundamentally different in the architecture; they're not simply off-the-shelf models with a token mask

markhammond68 - 9 hours ago

Great article, thanks!

bellajbadr - 17 hours ago

Is this only about getting fix json output?

vivzkestrel - 11 hours ago

- i have to keep scrolling down on your home page https://nishtahir.com/ to see what posts you have

- could you kindly put all that in a /blog page with pagination and not infinite scroll?

chelseahermes - 2 hours ago

[flagged]

fr2029 - 6 hours ago

[dead]

maidaman - 10 hours ago

[dead]

alexx-devv - 8 hours ago

[flagged]

hensenjuang - 11 hours ago

[flagged]

mehar_pro - 10 hours ago

[dead]