Skip to content
View original post on X: Kilo (acq. by Anaconda)· 20/100AI score20/100

Hands-on guide to writing evals that catch false agent claims

AISummary

A hands-on guide by @pandemicsyn walks through writing evals that detect when an AI agent claims to have completed a task it never did. Working through a demo agent that fails on purpose, the author refines the checks until they can distinguish real work from mere claims of work. The post includes a coding agent skill that can guide readers through the exercise.

Post on XView on X

An AI agent told the user "I filed your bug" - but nothing was filed, and the test still passed.

@pandemicsyn wrote a hands-on guide to evals, the tests that catch this kind of thing in AI agents. You work through a small demo agent that fails on purpose, and you fix the checks until they can tell the difference between an agent that did the work and one that only said it did. His coding agent skill can walk you through it if you'd rather not read the whole post.

If your test can't tell those two apart, you don't know what your agent is doing.

Read more here: https://neonronin.sh/blog/learn-to-write-evals/

Source: Kilo (acq. by Anaconda) · x.comPublished · added here