AWS debuts Strands Decider 2B, a first lightweight decision model for accelerate agentic workflows Amazon Web Services' Strands Labs team released Strands Decider 2B, an open-source decision model built on Qwen3.5-2B's base torso with a 1 million-parameter pointer head and a rank-16 LoRA adapter, available now via Hugging Face and GitHub. AWS said the 2 billion-parameter model runs locally with under 150 milliseconds of latency and targets agentic tasks including model routing, tool selection, context management, guardrail enforcement and policy classification. The release, labeled v.20, is AWS's first attempt at improving on decision models such as TypeSafe AI's Jev, which AWS said suffers performance issues on intricate reasoning due to its parallel output structure. AWS debuts Strands Decider 2B, a first lightweight decision model for accelerate agentic workflows Amazon Web Services Inc.’s https://aws.amazon.com/ Strands Labs team https://aws.amazon.com/blogs/opensource/introducing-strands-labs-get-hands-on-today-with-state-of-the-art-experimental-approaches-to-agentic-development/ has been playing around with an emerging class of lightweight artificial intelligence systems known as “decision models,” and it’s now making the fruits of that work available to the open-source community. The company said today it’s releasing an open-source decision model called Strands Decider 2B, which is designed to make decisions rapidly by eliminating the need to generate text, which eats up vast amounts of tokens and increases response latency. Strands Decider 2B has been optimized for local deployments and rapid experimentation, so developers can download and run it on their own laptops as well as public cloud environments. By giving developers access to a dedicated decision-making engine, AWS said, it hopes to accelerate the pace of agentic AI development more broadly. Decision models, also known as “System 1 models,” are very different things from the large language models that everyone knows and either loves or hates. Whereas LLMs respond to inputs by generating text, code, images or video, decision models simply make decisions when they’re presented with a number of predefined choices, with no other outputs. Each decision they make is assigned a confidence score to help users judge how accurate the response is likely to be. The omission of text generation means that decision models excel in terms of speed and low latency. The confidence scores assigned to each response are crucial, because the lack of text-generation capabilities means these models can’t explain their decisions. Though decision models have been around for a while, they only started getting attention a couple of weeks ago following the emergence of a startup called TypeSafe AI Inc. https://siliconangle.com/2026/09/16/typesafe-ai-exits-stealth-with-40m-to-build-ai-for-use-by-software/ and its subsequent release of Jev. It was designed https://typesafe.ai/blog/introducing-system-one-models-and-jev to make fast, structured decisions that software and AI agents can use directly. Jev demonstrated how useful decision models can be, but AWS’s team believes it suffers from a number of structural shortcomings. For one thing, its parallel output structure means there are performance issues when it’s asked to perform intricate reasoning. AWS decided to try to improve on the concept and Strands Decider 2B is its first real attempt at doing that. It’s built on top of a standard LLM, using Qwen3.5-2B’s base torso, but AWS swapped out the traditional “LLM head” for a tiny, customized “pointer head” that has a relatively minuscule 1 million parameters. It fine-tuned the Qwen3.5-2B torso using a rank-16 LoRA adapter to score hidden states of available choices directly against answer positions. The cloud giant said it has repeatedly iterated on Strands Decider 2B, with today’s release only named v.20. The company said it deliberately chose the 2 billion-paremater scale because it believes this size provides just the right balance, making it small enough to run on a local machine with less than 150 milliseconds of latency, yet still powerful enough to make complex decisions. Strands Decider 2B certainly stacks up well in benchmarks, with AWS reporting that it achieved high-level performance in terms of accuracy and calibration on JevBench when compared with other open-source 2B models. The model is available to download now via Hugging Face, and the full codebase, training scripts and examples can be found on its GitHub page. AWS hopes the AI developer community will experiment with Strands Decider 2B and accelerate agentic tasks such as model routing, tool selection, context management, guardrail enforcement and policy classification. It also paves the way for the creation of “hybrid agents” that use decision models to make simple and repetitive choices and LLMs for complex reasoning challenges. Image: SiliconANGLE/Gemini A message from John Furrier, co-founder of SiliconANGLE: Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network , where technology leaders connect, share intelligence and create opportunities. - 15M+ viewers of theCUBE videos , powering conversations across AI, cloud, cybersecurity and more - 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network Are you an AWS customer? Support SiliconANGLE financially by buying your AWS services from our Marketplace portal page and links: https://siliconangle.com/aws-marketplace/ https://siliconangle.com/aws-marketplace/ About SiliconANGLE Media SiliconANGLE https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fsiliconangle.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=SiliconANGLE&index=9&md5=646b1b564e2259100a2b8638aab0a552 , theCUBE Network https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecube.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Network&index=10&md5=7de2a85f95ab4a4a495cede20b8cb1da , theCUBE Research https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fthecuberesearch.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Research&index=11&md5=7bb33676722925eb57d588ec343e4f6f , CUBE365 https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.cube365.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=CUBE365&index=12&md5=d310fb35919714e66ad8d42c9c0c1bc6 , theCUBE AI https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecubeai.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+AI&index=13&md5=b8b98472f8071b23ebb10ab9a8dd0683 and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI. Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.