Architect launches Liquid Inference, a per-request auction router for LLM inference
Original titleArchitect Launches Liquid Inference, a Real-Time Auction for LLM Inference
AISummary
Architect Financial Technologies has launched Liquid Inference, an LLM router that auctions each request to providers quoting the requested model, and the lowest qualifying offer wins.
Buyers can set per-job cost caps, time-to-first-token limits, minimum throughput, and region or zero-data-retention rules, and the max price is locked before generation. The source states that fees, provider list, and latency data are not yet public.
Source: MarkTechPost · marktechpost.comPublished · added here