One of the more intriguing announcements at OpenAIâÂÂs Dev Day event on Tuesday came in an aside from CEO Sam Altman, who revealed the companyâÂÂs new âÂÂDecisions API.âÂÂ
The API apparently provides similar functionality to Jev, a model released by TypeSafe AI earlier this month thatâÂÂs explicitly designed for software automation. A kind of super-powered classifier built on an LLM, developers can give Jev a set of choices that it outputs as probabilities cheaply and at high speeds.
OpenAIâÂÂs Decisions API seems to be the same sort of product. At the event, Altman described the API as a way to give the labâÂÂs Luna model a predefined set of options to choose between, such as categories in which to classify an image or different agent behaviors.
âÂÂBy focusing the model on that choice, we can make it extremely fast while keeping capabilities like image understanding, broad language support, and safety protections,â Altman said.
TypeSafe didnâÂÂt respond to TechCrunchâÂÂs questions about the new product, but CEO Diogo Almeida, a former OpenAI engineer who co-invented reinforcement learning, joked on X about the beginning of the clone wars.
He added that OpenAIâÂÂs interest could be âÂÂa signâ¦that building in a System One compatible way is the future.â (âÂÂSystem Oneâ is TypeSafeâÂÂs term of art for fast, intuitive thinking, versus âÂÂSystem 2,â which it applies to deliberate reasoning.)
The subtext here is that LLMs as we know them arenâÂÂt the right solution for a lot of software because they are comparatively slow and expensive. Developers have been using Jev to augment LLMs and, in doing so, have found that theyâÂÂre faster and cheaper.
ItâÂÂs not clear how similar Decisions API will be to Jev, since OpenAI released it as a limited preview and, thus far, TechCrunch hasnâÂÂt spotted developers running it through its paces. However, there is clearly interest, according to the conversations on X.
Decisions API isnâÂÂt the only Jev-like API on the internetâÂÂother startups are rolling out similar models; OpenAI wonâÂÂt be the last tech giant to produce one. A key question is how well calibrated each of these decision modelsâ outputs will be to real life.
Almeida says his companyâÂÂs moat is the synthetic data it creates to generate statistically useful outputs.
âÂÂFast and cheap is very easy, you know,â Almeida told TechCrunch last week. âÂÂIf you want it really fast and cheap, use dice, right? Intelligence is the hard part, and my North Star is always pushing the intelligence-per-dollar Pareto curve.âÂÂ
After just weeks, it seems clear that these models have a future ahead of them, and one likely application is monitoring and securing AI agents. One of OpenAIâÂÂs new security measures following a series of incidents where its agents misbehaved on the open internet is using a separate model to watch for bad actions at âÂÂsignificant compute cost.âÂÂ
Shapor Naghibzadeh, a long-time cybersecurity professional who leads the start-up QueryStory, thinks that a model like Jev could make that possible far more cheaply.
He built a demo for a hackathon held last weekend that uses Jev to check each agentic action against the task it was given, blocking actions it had high confidence were bad, flagging others for review, and permitting the rest.
In theory, such monitoring could have stopped the Hugging Face incidentâÂÂand monitoring of that kind costs $2.94 with Jev, versus $372 with a frontier LLM.
A key observation is that Jev is arguably cheap enough to run on every agentic action, which offers a layer of review that could improve the reliability of agents writ large. ItâÂÂs the kind of thing TypeSafe was hoping to achieveâÂÂand now OpenAI has seen the value as well.
Topics
When you purchase through links in our articles, we may earn a small commission. This doesnâÂÂt affect our editorial independence.
