Skip to content
View original post on X: Design Arena· 40/100AI score40/100

GPT-6 Astra hedges far more than Claude Opus 5.5 in reasoning summaries

AISummary

Design Arena analyzed 324 thinking summaries and found OpenAI's GPT-6 Astra uses hedging words like "maybe," "might," and "it seems" about 20 times as often as Anthropic's Claude Opus 5.5. Opus usually weighs a few options and commits early, in about 4 out of 5 summaries versus 1 in 4 for Astra, which the post says works more like a designer while Opus works more like a builder.

Post on XView on X
@DesignArena

@OpenAI's GPT-6 Astra hedges ("maybe", "might", "it seems") about 20x as often as @AnthropicAI's Claude Opus 5.5. Opus usually weighs a few options and commits early, in about 4 out of 5 summaries as opposed to Astra's 1 in 4.
We read 324 of their thinking summaries on Design Arena to see how each one actually goes about building a game. Astra works more like a designer, Opus more like a builder.
Full video in the replies.

Source: Design Arena · x.comPublished · added here