OpenAI’s Jev clone could help the frontier lab stop its swarming agents

Nairavoice | 1h ago 110 0 1 min read
OpenAI’s Jev clone could help the frontier lab stop its swarming agents

After just weeks, it seems clear that these models have a future ahead of them, and one likely application is monitoring and securing AI agents. One of OpenAI’s new security measures following a series of incidents where its agents misbehaved on the open internet is using a separate model to watch for bad actions at “significant compute cost.”

Shapor Naghibzadeh, a long-time cybersecurity professional who leads the startup QueryStory, thinks that a model like Jev could make that possible far more cheaply.

He built a demo for a hackathon held last weekend that uses Jev to check each agentic action against the task it was given, blocking actions it had high confidence were bad, flagging others for review, and permitting the rest.

In theory, such monitoring could have stopped the Hugging Face incident — and monitoring of that kind costs $2.94 with Jev, versus $372 with a frontier LLM.

A key observation is that Jev is arguably cheap enough to run on every agentic action, which offers a layer of review that could improve the reliability of agents writ large. It’s the kind of thing TypeSafe was hoping to achieve — and now OpenAI has seen the value as well.

❤Enjoying this article? Support our work with a small crypto donation.
Show Some Love By Sharing

Discover more from NAIRAVOICE.COM.NG

Subscribe to get the latest posts sent to your email.

❤Enjoyed this? A small crypto donation helps us keep publishing.
Nairavoice
Nairavoice

Contributor at NairaVoice.com.ng

Related Posts

Leave a Reply