Strands Decider 2B: a small, open-source, decision model

strandsagents.com

274 points by gmays 18 hours ago


real_faxenoff - 13 hours ago

As a regular user of a bunch of specialized micromodels, I'll tell you this: you won't be happy with such a model (and its JEV counterparts) running permanently in the background on your PC's CPU. You need to offload their processing to the NPU. There are many pitfalls along the way, but the result is worth it.

NPU performance will be twice as high, while power consumption will be four times lower. No additional fan noise (if you know what I mean).

I'll wait another month until the first phase of the =battle royale= among models of this kind wraps up, put together a solution for the NPU/iGPU, and post it on HF.

adenta - 17 hours ago

At this point I can't wait for a comedian to release a decision model backed by humans.

Meet Jerry- it's literally a guy named Jerry answering your questions.

mattvr - 14 hours ago

Why is everyone calling binary choices `noul`? Does this have some meaning or is it just copying Jev’s API?

keyle - 17 hours ago

Fantastically well written. It's rare for me to be able to understand what the AI gurus are talking about, and this was written by humans for humans.

It can technically be used for a lot of use cases, I'd like people to chime in on ideas on this?

gsnedders - 5 hours ago

What I don’t get from this article is why strands-decider-2b v10 and strands-decider-2b v11 are shown as “other systems”, and with v11 notably outperforming v19 — why are these other systems, and why is v11 not the path taken?

woadwarrior01 - 13 hours ago

The ~2-week-old Intern-Decision family of models (0.8B, 2B and 4B) have the same Qwen3.5 base model family (albeit the instruction-tuned variants) and pointer head architecture.

https://huggingface.co/collections/internlm/intern-decision

miguelspizza - 18 hours ago

This is a great model. I've been running it on device in chrome extension to filter things like email.

It is just the right mix of size, capability and speed to make it generally useful for adhoc bulk classification tasks.

For those wanting to run it in browser: https://huggingface.co/alxnahas/strands-decider-2B-webgpu

hrpnk - 14 hours ago

clef from cloudflare runs on llama.cpp - being locked-in to strands cli would be a bummer and will slow down adoption.

Since it's a LoRa on Qwen, I assume this is runnable via llama.cpp. Pity that the PEFT/LoRa->GGUF translation is left to the user. Anyone got past:

    $ uv run --with transformers==5.19.0 convert_lora_to_gguf.py ~/Downloads/lora --dry-run --verbose
    [...]
      File "/Users/user/repos/llama.cpp/conversion/base.py", line 630, in map_tensor_name
    raise ValueError(f"Can not map tensor {name!r}")
    ValueError: Can not map tensor 'layers.0.linear_attn.in_proj_a.weight'
dev_l1x_be - 7 hours ago

General question: what is this model good for? I have mixed results with Laya on a pdf classifier (Jev was doing much better).

ricardobeat - 9 hours ago

I wish these would stop using JevBench. It focuses way too much on text classification tasks, and some of the models perform very poorly on tasks that need actual intelligence.

girvo - 12 hours ago

Does anyone know if it is worth fine-tuning one of these decision models on the shape of the questions you want it to work on, vs the more general versions? I'm using Jev pretty successfully at work at the moment, but am curious about what is doable

SubiculumCode - 15 hours ago

Are any of these multimodal yet? I'd love to try asking a model with calibrated probabilities to answer question like, "do these shapes match?". Sure, you can ask a LLM....

mynti - 13 hours ago

Can someone explain this architecture a bit more in depth? They say the pointer head scores the hidden state at each option against the hidden state of the answer. But the LLM produces hidden states per token, so an option can span multiple tokens, no?

AdamIdrissi - 3 hours ago

Wow this is rlly cool.

teruakohatu - 16 hours ago

Any idea how well this would run on a CPU?

soltanov - 15 hours ago

Benchmark calibration does not establish reliability on unfamiliar production inputs.

kimseungyong - 11 hours ago

I wait for this open-source model.

I should let the development agent select a model and run it to improve token efficiency.

Is it being used this much these days?

mijoharas - 7 hours ago

so, how are people running these locally? is there a llama.cpp/ollama solution for systemone/jev style apis?

CMay - 3 hours ago

The kind of stuff that people are doing with decision models today is largely the way I've been using Liquid AI's LFM models, as rapid iteration interrogation tools for atomic decisions. They aren't good at reasoning, but you can reason about what decisions you would need in order to come to a reasonably high quality decision about a thing.

Even though they weren't themselves decision models exactly, they were so fast that you could use them in a similar way. They obliterated Qwen and other models for a lot of tasks in that use case.

I'm not sure I would agree with some of the claims made in this article though. You can definitely do a lot of similar tasks to regular LLMs with decision models, you just have to chain the results. You can even technically infuse a broader perspective into each token choice or force certain context to have priority in the decision making of the next token.

ContinuityLab - 10 hours ago

Swapping out the text-generation head for a dedicated pointer head on a small footprint model is a pragmatic approach for low-latency local decision pipelines.

- 7 hours ago
[deleted]
muslimpeacepriz - 3 hours ago

Gone are the days when HN was flooded by "-lang.org"-postw; nowadays it is " model". Equally full of BS.

davvie - 16 hours ago

Looks really nice, I think I could use it on my Mac mini for some smaller automations

stephantul - 14 hours ago

2B being called small is such a sign of the times

Ujj-001 - 12 hours ago

is jev commoditized now ?

cosprax - 4 hours ago

[flagged]

chelseahermes - 6 hours ago

[flagged]

lin7c - 14 hours ago

[flagged]

happybox2016 - 11 hours ago

[flagged]

Fluid_Mechanics - 17 hours ago

[flagged]

yieldcrv - 15 hours ago

a strand type game