Skip to content
View original post on X: Ethan MollickX· 43/100AI score43/100

Opus 5 also beats Montezuma's Revenge, and Metaculus recreates a test

AISummary

Ethan Mollick says Anthropic's Opus 5 beat Montezuma's Revenge, as well as OpenAI's GPT-6 Astra. He notes that one criterion in the AGI bet is for AI to win a discontinued, weak Turing-style prize, and Metaculus has decided to recreate that test to confirm whether the criterion is resolved.

Post on XView on X
Ethan MollickVerified on X
@emollick

Not only did Astra beat Montezuma's Revenge, but so did Opus 5

But one of the criteria to resolve this AGI bet is for AI to win a discontinued prize that was sort of like a weak Turing Test. So Metaculus has decided to recreate the test to confirm the criteria is resolved. Neat!

Metaculus@metaculus
Last month GPT-6 Astra (@chatgpt) beat Montezuma's Revenge, and a lot of people declared that our "weakly general AI" question, originally launched in 2020, resolved. But…it still hasn’t resolved. We wanted to give everyone a quick update on where things actually stand and the steps we’re currently taking (cc @swishfever @joeykrug @emollick). 🧵
View quoted post on X

Source: Ethan Mollick · x.comPublished · added here