Technology
AWS open-sources Strands Decider 2B for fast local agent decisions
The 2-billion-parameter model, built on Qwen3.5-2B, scores predefined options with confidence instead of writing free text; weights are on Hugging Face under Apache 2.0.
Published: October 2, 2026 · 1 min read
Amazon Web Services released Strands Decider 2B on Oct. 1, an open-source “decision model” meant to choose among predefined options in tens of milliseconds — a cheaper, more structured step for agent workflows than calling a full large language model every time.
Announced on the Strands Agents blog by AWS distinguished engineer Marc Brooker and colleagues, the roughly two-billion-parameter system starts from Alibaba’s Qwen3.5-2B torso, strips the text-generation head and adds a small scoring head so it returns calibrated choices and confidence scores in one forward pass. Weights are on Hugging Face and training code is on GitHub under Apache 2.0; AWS is not offering a hosted pay-per-call API. Brooker told TechCrunch the idea grew from customer agent pipelines that needed reliable next-step routing — tool selection, guardrails, model routing — without frontier-model latency or cost. On a local Nvidia RTX 3090, AWS reports median decision latency around 115 milliseconds for small tasks; an M3 MacBook landed near 153 milliseconds in their tests.
Decision models surged after TypeSafe’s Jev drew attention earlier this year. TechCrunch noted Amazon’s release landed the same week OpenAI floated a similar idea, while TypeSafe CEO Diogo Almeida argued many clones still lag on real intelligence. Strands Labs positions Decider 2B as an experiment-friendly baseline: small enough to fine-tune on ordinary hardware, competitive on the public JevBench leaderboard for its size class, and deliberately unsuitable for chat, coding or long-form generation.
For Canadian developers and enterprises already wiring agent stacks, a free local decider lowers the cost of putting a confidence-gated checkpoint before tool calls — provided teams accept the ops burden of self-hosting and treat scores as workflow signals, not guarantees.
Sources: Strands Agents / AWS Strands Labs; TechCrunch.
Sources
- Strands Agents / AWS Strands Labs · other
- TechCrunch · other
Newsletter
News. Context. What matters.
One essential briefing, written for people who would rather understand the story than scroll it.
Unsubscribe anytime. We don’t sell addresses.
Recommended
Technology
Micrologic launches Canadian Sovereign Digital Vault for cyber-recovery data
Unveiled at Québec City’s Convergence 2026, the isolated Canadian-jurisdiction copy aims to help organizations restart after ransomware; $45-million from FTQ and Investissement Québec backs the build.
Technology
SpaceX lofted Google TPUs to orbit in first Project Suncatcher AI test
Transporter-18 carried a Planet Labs solar-powered prototype with Alphabet’s tensor processors — an early check on whether space can host future machine-learning infrastructure.
Technology
OpenAI says it disrupted Moonshot-linked campaign to extract model reasoning
The company attributes a core cluster of “adversarial distillation” traffic to people associated with China’s Moonshot AI, developer of Kimi, after spikes of thousands of extraction-pattern requests in July.