Introducing Liquid Inference by Architect, a new LLM model router where inference providers compete dynamically to offer lowest-cost tokens. Providers onboard in minutes, not weeks. Buyers have full price visibility and receive 20% of referred fees in free inference. More info:
Liquid Inference can be used in place of any OpenAI or Anthropic API-compatible backend. Behind it, our team has leveraged our experience trading and building financial exchanges to create real-time two-sided price discovery for inference. Benefits for buyers and providers:
INFERENCE BUYERS
• Fully compatible with agentic coding tools: Claude Code, Codex, OpenCode, Cursor, Pi, Cline, and many more. Full multi-modal support.
• Hundreds of open- and closed-weight models. Create routing rule presets or use our Auto routing algorithms.
• Providers compete dynamically on price for every prompt, offering the lowest marginal cost.
• Sign up for free with email, first 500 users get $20 of free inference.
• Receive 20% of referred fees as free inference, 10% for second-level referrals.
INFERENCE PROVIDERS
• Onboard through the Liquid Inference app. We verify new providers in minutes, not weeks.
• All prompts are OpenAI API standard. Simple REST/Websocket API to register models and quotes.
• Update quotes dynamically based on your own costs. Monetize GPU capacity only when you want.
• Standardize payouts via Stripe, full itemized records of all jobs performed.