Skip to content
View original post on X: TinkerOfficial· 32/100AI score32/100

Tinker removes extra prefill charges for 128k and 256k context

AISummary

Tinker says prefill for 128k and 256k context no longer costs extra, an effective discount of over 2x for long-context models including Kimi K2.6, gpt-oss-120b, and Inkling. Prefill is also discounted for seven Qwen and Nemotron models, and sampling is cut for Qwen3.5-9B and 9B-Base.

Post on XView on X
TinkerVerified on X
@tinkerapi

Part of a thread · earlier post

Prefill for 128k and 256k context no longer costs extra, an effective discount of over 2x for long-context models such as Kimi K2.6, gpt-oss-120b, and Inkling.

Prefill has been further discounted for 7 models from Qwen and Nemotron, and sampling for Qwen3.5-9B and 9B-Base.

Source: Tinker · x.comPublished · added here