OpenAI's Ultrafast Astra runs up to 8x faster in Codex
Original titleExcited to see what new bottlenecks people complain about next when model inference is ~basically instant 🙃
AISummary
OpenAI's Ultrafast mode for Astra runs up to 8x faster than Astra Standard and 4x faster than Astra Fast in Codex. Sherwin Wu says inference now feels nearly instant, shifting the bottleneck to users' next complaints.
Source: Sherwin Wu · x.comPublished · added here