Skip to content
View original post on X: Harrison ChaseX· 36/100AI score36/100

Harrison Chase highlights Jev, a small model for agent yes/no decisions

AISummary

Harrison Chase shares a post from @ch3nweiii arguing that many agent steps are yes/no calls rather than generation. He points to Jev, which answers those typed questions with calibrated probabilities so the frontier model handles only the hard work, and links a LangChain blog post on building a harness with Jev.

Post on XView on X
Harrison ChaseVerified on X
@hwchase17

good framing from @ch3nweiii - a lot of agent steps are yes/no calls, not generation

Jev answers those as typed questions with calibrated probabilities, so the frontier model sticks to the hard work

where those calls sit in a harness: https://www.langchain.com/blog/building-a-harness-with-jev

https://x.com/ch3nweiii/status/2108664035109470637

Chen@ch3nweiii
whoever built this realized we've been using our smartest AI for the dumbest possible jobs your $20-$200/mo frontier model is researching, writing code, planning… and then you're paying that same brain to decide YES / NO. Jev flips that. it's a tiny decision model that sits in front of the expensive stuff and takes the boring forks: -> which agent goes next? -> is this worth researching? -> reply or ignore? -> publish, wait or escalate? -> BUY / SELL / HOLD? -> click this or keep looking? then I saw the numbers people are posting and the whole architecture started making a lot more sense. ~1,000 papers classified for around $0.08 ~500 emails sorted for around $0.035 browser-agent loops shown at around 7 seconds for ~$0.0039 and at the quoted pricing, roughly 10,000 ~1K-token decisions comes out around $0.42. that's not because Jev got smarter than Claude. it's because nobody asked it to be Claude. the setup is stupidly simple: -> big LLM gets the hard thinking, research and writing -> Jev gets the thousands of tiny decisions between those steps -> code actually clicks, saves, sends and runs and suddenly you start seeing these little decisions everywhere. > your inbox. > X feed. > support queue. > meetings. > leads. > video clips. > browser agents. > even paper-trading loops. that's the part I think most agent builders are missing. the expensive model doesn't need to touch every step just because it's able to touch every step. > let Claude think. > let Jev decide what deserves Claude. > let code do the actual work. I broke down how to wire Jev into your agent + the exact router setup in the article below
View quoted post on X

Source: Harrison Chase · x.comPublished · added here