Image source: Public Domain
Architect Financial Technologies Inc. ("Architect" or "the Company") announced the launch of Architect Liquid Inference ("Liquid Inference"), a live auction platform for large language model (LLM) inference. On Liquid Inference, inference providers post competing offers to serve AI models, and every request is awarded to the lowest-priced offer that satisfies the buyer's rules for cost, intelligence, latency, and data handling.
Liquid Inference applies market structure principles from regulated derivatives trading to AI inference, while offering a chat-style interface and standard LLM API to users. The platform publishes live order books, per-provider and per-model quotes, and cleared transactions for account holders. Buyers can attach purchasing rules to any request, including per-job cost caps, maximum time to first token, minimum throughput, approved regions, zero data retention, and provider or model allow lists. Liquid Inference also supports auto-routing features to dynamically choose the optimal model for a particular unit of work.
Liquid Inference is compatible with all existing AI code tools and agentic workflows. Buyers can use any OpenAI- or Anthropic-compatible client, including the official SDKs and coding tools such as Claude Code, Codex, Cursor, and Cline. The platform auctions the requests across every provider quoting the named model, locks a maximum price before the first token is generated, and charges only on metered usage. Each job produces a fixed receipt showing the winning provider, the price cap, the final charge, and the depth of competition at the time of award.
Brett Harrison, Founder and Chief Executive Officer of Architect, commented on the launch, "Inference is quickly becoming one of the largest line items in every technology budget, but it is still only available via static list prices and bilateral contracts. Liquid Inference brings the core tools of modern exchanges to every AI request: competitive quoting, transparent market data, and firm limit prices. Developers get the best available price without changing their code, and providers get a neutral venue to compete on cost and performance. Providers also have a place to monetize underutilized GPU clusters."
The launch extends Architect's wide-ranging enterprise to financialize AI inputs. The Company is operationalizing the American Innovation Exchange ("AI Exchange"), a U.S. Designated Contract Market acquired in May 2026, as the first CFTC futures exchange purpose-built for GPU compute futures and options and AI supply chain commodities pending regulatory review. Architect additionally brokers block transactions in compute capacity and structures OTC swaps on compute and memory through its CFTC Introducing Swaps Broker subsidiary Architect Financial Derivatives LLC. In launching Liquid Inference, Architect adds a live spot market for inference to its extensive suite of AI trading and financial services.