AWS open-sources Strands Decider, a 2B model that picks an answer and says how sure it is

Wednesday, 7 October 2026By Sridhar Mukkandi1 min read

Strands Labs at AWS has released Strands Decider, an open decision model under the Apache 2.0 licence. It takes Qwen3.5-2B-Base, removes the part that generates text, and adds a small pointer head of about a million parameters that scores the given options in one pass. The model has 1.9 billion parameters, answers in a median 115 ms on an RTX 3090, and also runs on Apple-silicon Macs and on CPU.

It answers three kinds of question: yes or no, pick one of several options, and rate on a scale. The team reports 76.2% accuracy (176 of 231 tasks) on the public set of JevBench, a third-party benchmark for decision models. It also says that on short tasks the model has not seen, answers given with a confidence of 0.9 or more are right about 95% of the time.

So what

A small model that runs on a laptop and reports its own confidence is a cheap first gate for agent steps such as routing, tool choice and guardrails. The team's own guidance is the part to copy: below 0.9, confirm the answer or ask a person.

SourceStrands Labs on GitHubgithub.com/strands-labs/strands-decider Go deeper · EssayWhy 0.5 is almost always the wrong threshold

Drafted with AI from the original source, then checked and edited by Sridhar Mukkandi. Spotted a mistake? Write to hello@tarkika.com and we'll correct it here.

More from The Signal

Models Cloudflare releases Clef, open decision models that answer with probabilitiesThe probability is what makes a decision model useful: code can act on a confident answer and hand a borderline one to a person. Check how well those probabilities match reality on your own data before you pick the cut-off. Cloudflare blog Models Reflection announces Beam, a 501B open-weight model it says needs 3 to 4 times less computeIf the efficiency claim holds, an open model near the top of the open field at a third of the compute would cut the cost of serving reasoning-heavy agents. Wait for the weights and run your own tests before planning around it. Reflection AI Careers Anthropic commits $100 million to train 10,000 engineers to deploy ClaudeThe scarce skill is no longer calling a model; it is getting one working inside a real organisation. Anthropic says it wants engineers with a record of building with LLMs and of helping others adopt AI, so shipped, adopted projects count for more than certificates. AnthropicAll stories