Skip to content
View original post on X: InferactOfficial· 13/100AI score13/100

Inferact runs open models in production, upstreaming vLLM tuning

AISummary

Inferact, founded by the creators and core maintainers of vLLM, says its platform runs open models in production on customer compute or its own. The company says changes for hardware such as NVIDIA Vera Rubin are landing in open-source vLLM and that the platform benefits directly from that upstream tuning.

Post on XView on X
InferactVerified on X
@inferact

Part of a thread · earlier post

These changes land in open-source vLLM, and the Inferact platform benefits directly from that upstream tuning. Founded by the creators and core maintainers of vLLM, Inferact runs open models in production on your compute or ours.

Early work on hardware like NVIDIA Vera Rubin shows us where inference is heading, and we use what we learn to optimize inference for your workload.

Learn more at http://inferact.ai

Source: Inferact · x.comPublished · added here