Frontier and open models. One API. Priced below list.
Everything about price, in one place — no scrolling required.
gpt-6.1-solclaude-opus-5-5claude-fable-5-1gpt-6-astraclaude-sonnet-5-5grok-4.7deepseek-v4.1-flashgpt-6-solgpt-6-lunaqwen3.8-max-primeglm-5.3-primemimo-v2.6-proqwen3.8-flash-nextqwen3.8-omni-flashnemotron-3.5-lightningclaude-opus-5gpt-5.6-solgpt-5.6-terragpt-5.6-lunaclaude-fable-5claude-opus-4-8claude-sonnet-5claude-haiku-4-5jev-1.13gemini-3-pro-imagePrices are exact per-token rates, not rounded — some carry more decimals than others. Each struck-through figure is that row's reference list price and links to the page that publishes it, so every line can be checked. "At list" means we sell at the reference price; "mixed" means we are below it on one side and above on the other. Every row checkable, every model verifiable →
Building an agent? Fund a key programmatically with USDC — no account, no card. x402 docs →
Plug in your monthly usage to see what it costs here, next to published list rates.
Side-by-side pricing vs every competitor
qspBuilt for terminals and AI agents. --json output with stable exit codes — Claude Code, Cursor, Aider can call it without parsing HTML.
1# Two lines. That is the whole migration.2from openai import OpenAI34client = OpenAI(5 base_url="https://api.quicksilverpro.io/v1",6 api_key="your-api-key",7)
Common questions
QuickSilver Pro is an OpenAI-compatible inference API. The current catalog is DeepSeek V4 Flash, DeepSeek V4.1 Flash, DeepSeek V4 Pro, Qwen3.8 Max, Qwen3.8 Max Prime, Qwen3.8 Omni Flash, Qwen3.7 Max, Qwen3.7 Plus, Qwen3.7 Flash, Qwen3.6 Plus, Qwen3.6-35B-A3B, Qwen3.8 27B, Qwen3.8 Flash Next, Kimi K2.6, Kimi K2.7 Code, Kimi K3, Muse Spark 1.3, Muse Spark 1.2, Muse Glimmer 30B, GLM 5.3, GLM 5.3 Prime, GLM 5.3 Flash, GLM 5.2, Nemotron 3.5 Lightning, GPT-OSS 120B, GPT-6 Astra, GPT-6 Luna, GPT-6.1 Sol, GPT-6 Sol, GPT-5.6 Luna, GPT-5.6 Terra, GPT-5.6 Sol, Grok 4.5, Grok 4.6, Grok 4.7, MiniMax M3, MiMo-V2.6-Pro, MiMo-V2.6-Flash, MiMo-V2.5, Hy4 Preview, Hy3, Jev 1.13, Claude Opus 5.5, Claude Opus 5, Claude Fable 5.1, Claude Fable 5, Claude Opus 4.8, Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.1 Pro Preview, Gemini 3 Pro Image, Gemini 3 Flash Preview, Gemini 3.1 Flash Lite, FLUX.2 Pro, FLUX.1 Schnell, SDXL Turbo, FLUX.2 Klein, Qwen-Image Max, Seedream 5.0 Pro, Seedream 4 and Bria FIBO 1.5, served through one endpoint and one API key.
V4 Flash revision 0731 is the current agent build. It has 1M context and supports Chat Completions and Responses.
Up to 20% below the standard published per-token list rate on most of the catalog, and 50–67% below on Claude. DeepSeek V4 Flash: $0.086 / $0.173. DeepSeek V4 Pro: $0.70 / $2.10. Kimi K3: $2.55 / $12.75. GLM 5.2: $1.12 / $3.52. Claude Opus 5: $2.00 / $10.00. GPT-5.6 Sol: $1.60 / $8.00. Grok 4.5: $1.60 / $4.80. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-5.6, Gemini and Grok included.
Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, and usage.cost accounting all work out of the box. json_schema strict mode is model-dependent: the Claude models do not support it, so a schema there is advisory and the enforced path is a tool with strict: true.
Yes, two ways. Every account gets a free trial balance to use in the browser chat at /chat — no card, no API key. And your first credit purchase is matched 100%, up to $50 in bonus credits. The bonus is added to your balance the next time you top up (any amount from $5): pay $20, then top up $5 later, and $20 of bonus arrives with it. A larger first payment still receives the $50 maximum. One per customer and payment card; standard pay-as-you-go after that.