跳到正文
原文
TechCrunch · AI· Tim Fernholz·· 5 小时前AI 评分52

AWS发布开源决策模型Strands Decider 2B

Amazon releases its own Jev clone as decision models flood the web

AI 导读

Amazon Web Services发布开源决策模型Strands Decider 2B,以高速、低成本方式在预先决定的选项中做选择并给出置信度,模型现已可用且小到可本地运行。

正文

Amazon Web Services released an open-source decision model inspired by TypeSafe’s Jev, with AI developers increasingly seeking intelligence that is more suited to computer automation than frontier LLMs.

Amazon’s Strands Decider 2B, released the same week OpenAI announced a similar offering, is a high-speed, low-cost way to sort between pre-decided options and deliver a measure of how confident it is in its choice. The model is fully open-sourced, available now, and small enough to run locally.

Amazon distinguished engineer Marc Brooker came up with the project after seeing Jev and trying to build his own take on such a model. The homebrew project was successful enough—it briefly reached the top spot on the Jevbench ranking for models of its size—that Amazon engineers cleaned it up and released it as an offering from their Strands Labs, an organization developing new tools and protocols for deploying AI agents.

Brooker says the need for a tool like this emerged in conversations with AWS customers, whose agentic workflows didn’t always require the capability or cost of a fully-featured LLM all the time.

“What originally piqued my interest in this class of models was that they make a perfect decider for a workflow step— ‘what is the next thing for me to do here, based on where I am?’” Brooker told TechCrunch. He said it offers customers “a workflow step that can be structured in a way that is more reliable, thanks to the confidence scores, thanks to the closed domain of answers, [and is] lower latency, potentially lower cost.”

Like other decision models, Strands Decider is built on the “torso” of an LLM, in this case Qen3.5-2B, but instead of generating text, it delivers calibrated choices. TypeSafe named their model Jev after the economist William Stanley Jevons, with hopes of invoking his theory that the falling cost of something—like computer intelligence—can, in fact, increase its demand.

The fact that dozens of similar models have been produced by researchers since TypeSafe debuted its idea shows the wide interest, but also raises the question of how valuable they can be. Brooker suggests that the challenge will be in optimizing the model’s speedy decision-making without compromising its intelligence.

“There is a very careful balance to be found where you want to push its performance on accuracy and calibration on these kinds of tasks, without degrading its performance on understanding different languages, on having the kind of knowledge it has, which is what makes it general purpose and interesting and useful,” he told TechCrunch.

Still, he doesn’t necessarily expect the frontier labs to dominate the space, especially since, with smaller markets, the cost to build something interesting is in the hundreds or thousands of dollars.

For their part, TypeSafe executives say they are keeping their heads down and improving future models.

“I get that people think it’s a gold rush, but they might be underestimating the difficulty of making the models actually smart,” CEO and founder Diogo Almeida told TechCrunch, saying that for now, he didn’t see real competition for his company emerging yet.

“The current batch seems more like ML people wanting to implement a cool architecture than a team deeply dedicated to making intelligence useful.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Tim Fernholz is a journalist who writes about technology, finance and public policy. He has closely covered the rise of the private space industry and is the author of Rocket Billionaires: Elon Musk, Jeff Bezos and the New Space Race. Formerly, he was a senior reporter at Quartz, the global business news site, for more than a decade, and began his career as a political reporter in Washington, D.C. You can contact or verify outreach from Tim by emailing tim.fernholz@techcrunch.com or via an encrypted message to tim_fernholz.21 on Signal.

View Bio

来源:TechCrunch · AI · techcrunch.com