Prefill for 128k and 256k context no longer costs extra, an effective discount of over 2x for long-context models such as Kimi K2.6, gpt-oss-120b, and Inkling.
Prefill has been further discounted for 7 models from Qwen and Nemotron, and sampling for Qwen3.5-9B and 9B-Base.
