Skip to content
Read the original: Cognition Blog (Devin, Windsurf)· 39/100AI score39/100

FrontierCode 1.1 refines its code-quality benchmark to curb unfair internet use

Original titleFrontierCode 1.1

AISummary

Cognition released FrontierCode 1.1, an update to its code-quality benchmark that adds a fair internet use prompt and a verifier that zeroes out runs consulting upstream fixes. The company also relaxed 75 of over 1,000 grading criteria, added scores for Sonnet 5 and updated scores for Fable 5, and dropped reporting on the Diamond subset.

Read the original cognition.com

Source: Cognition Blog (Devin, Windsurf) · cognition.comPublished · added here