Skip to content
Read the original: MarkTechPost· Published 60/100AI score60/100

Architect launches Liquid Inference, a per-request auction router for LLM inference

Original titleArchitect Launches Liquid Inference, a Real-Time Auction for LLM Inference

AISummary

Architect Financial Technologies has launched Liquid Inference, an LLM router that auctions each request to providers quoting the requested model, and the lowest qualifying offer wins.

Buyers can set per-job cost caps, time-to-first-token limits, minimum throughput, and region or zero-data-retention rules, and the max price is locked before generation. The source states that fees, provider list, and latency data are not yet public.

Read the original marktechpost.com

Source: MarkTechPost · marktechpost.comPublished · added here