Architect Financial Technologies has rolled out Liquid Inference, an LLM router that runs a live auction for every request. Providers submit bids to serve each prompt, and the buyer pays the lowest offer that meets its rules. Developers only need to change the base URL in their code to use the service.
Why it matters
Developers can potentially lower inference costs by automatically selecting the cheapest provider that meets their criteria.