Pricing

Frontier and open models. One API. Priced below list.

Everything about price, in one place — no scrolling required.

Sort
Filter
Model
Context
Latency
Intelligence
Input
Output
vs. list
GPT-6.1 SolNew
gpt-6.1-sol
GPT-6.1 Sol: near-Astra agentic coding and computer use at Sol pricing
ReasoningToolsVision
1M
$2.00 ($4.00 >200K)
$10.00 ($15.00 >200K)
claude-opus-5-5
most capable Claude: agentic coding, knowledge work, computer use
ReasoningToolsVision
1M
58
$1.60$4.00
$8.00$20.00
−60%
Claude Fable 5.1New
claude-fable-5-1
deep reasoning, long-horizon agentic (5.1)
ReasoningToolsVision
1M
57
$4.00$10.00
$20.00$50.00
−60%
gpt-6-astra
long-horizon agentic coding, deep research, document work, image input
ReasoningToolsVision
1M
55
$10.00 ($20.00 >200K)
$50.00 ($75.00 >200K)
Claude Sonnet 5.5New
claude-sonnet-5-5
newest Sonnet: faster, fewer tokens per task, strong coding
ReasoningToolsVision
1M
$1.00$2.00
$5.00$10.00
−50%
gemini-3.8-flash
most capable Flash for agents; 1M context
ReasoningToolsVision
1M
47
$0.6375$0.75
$3.1875$3.75
−15%
grok-4.7
long-running agentic coding, knowledge work, self-verification, image input
ReasoningToolsVision
500K
$2.00 ($4.00 >200K)
$6.00 ($12.00 >200K)
qwen3.8-max
Qwen 3.8 flagship, autonomous coding
ReasoningToolsVision
1M
47
$2.00
$6.00
glm-5.3
complex software engineering, long-horizon agents
ReasoningTools
1M
49
$1.12$1.40
$3.52$4.40
−20%
deepseek-v4.1-flash
newest DeepSeek Flash architecture, faster than V4 Flash
ReasoningTools
1M
$0.125$0.30
$0.55$1.20
−54%
kimi-k3
Multimodal reasoning, agentic coding
ReasoningToolsVision
1M
50
$2.55$3.00
$12.75$15.00
−15%
GPT-6 SolNew
gpt-6-sol
GPT-6 mid-flagship: strong reasoning, agentic coding, a rung below Astra
ReasoningToolsVision
1M
$2.00 ($4.00 >200K)
$10.00 ($15.00 >200K)
GPT-6 LunaNew
gpt-6-luna
cheapest GPT-6 tier: high-volume chat, classification, lightweight agents
ReasoningToolsVision
1M
$0.10 ($0.20 >200K)
$0.50 ($0.75 >200K)
deepseek-v4-pro
premium reasoning
ReasoningTools
1M
42
$0.70$1.32
$2.10$3.96
−47%
muse-spark-1.3
Flagship coding & agentic reasoning, 1M context
ReasoningToolsVision
1M
53
$1.00$1.25
$3.40$4.25
−20%
minimax-m3
long-horizon agentic coding, tool use
ReasoningTools
1M
36
$0.24$0.30
$0.96$1.20
−20%
Qwen3.8 Max PrimeNew
qwen3.8-max-prime
Qwen top-tier flagship: deepest reasoning, autonomous coding
ReasoningToolsVision
1M
$4.00
$12.00
GLM 5.3 PrimeNew
glm-5.3-prime
top reasoning flagship, complex software engineering
ReasoningTools
1M
$2.24$2.80
$7.04$8.80
−20%
mimo-v2.6-pro
flagship open-weight coding, agentic workflows, image input
ReasoningToolsVision
1M
46
$0.348$0.435
$0.696$0.87
−20%
glm-5.3-flash
high-volume agents and coding at workhorse pricing
ReasoningTools
1M
46
$0.06$0.075
$0.20$0.25
−20%
deepseek-v4-flash
fast chat & coding, thinking on by default
ReasoningTools
1M
41
$0.086
$0.173
gpt-oss-120b
high-volume production agents, OpenAI open-weight MoE
ReasoningTools
131,072
16
$0.041$0.15
$0.187$0.60
−69%
qwen3.8-27b
1M-context coding agents at small-model prices
Tools
1M
41
$0.34$0.425
$2.04$2.55
−20%
qwen3.8-flash-next
Cheapest next-gen Qwen: 6B-active Qwen4 preview on a 1M-token window
Tools
1M
$0.12$0.15
$0.376$0.47
−20%
qwen3.8-omni-flash
Qwen's first agentic omni model: image, audio + video understanding with tool use, at flash pricing
ReasoningToolsVision
1M
$0.15
$0.47
glm-5.2
long-horizon agents, project-level coding
ReasoningTools
1M
$1.12$1.40
$3.52$4.40
−20%
nemotron-3.5-lightning
fast, tool-heavy agents and high-volume automation
ReasoningTools
262K
$0.066
$0.176
qwen3.7-max
Qwen 3.7 flagship, agent / coding
ReasoningTools
1M
$1.25$1.475
$3.75$4.425
−15%
qwen3.7-flash
fast multimodal agents, visual coding, search
ReasoningToolsVision
1M
$0.024$0.03
$0.104$0.13
−20%
kimi-k2.6
Opus-class agentic / planning
ReasoningToolsVision
256K
$0.5472
$2.728
muse-spark-1.2
Coding-focused reasoning, agentic dev
ReasoningToolsVision
1M
47
$1.00$1.25
$3.40$4.25
−20%
muse-glimmer-30b
Budget agents, everyday coding, open weights
ReasoningToolsVision
131,072
24
$0.28$0.35
$1.20$1.50
−20%
Claude Opus 5
claude-opus-5
demanding reasoning, end-to-end coding, visual analysis, long-horizon agents
ReasoningToolsVision
1M
54
$2.00$5.00
$10.00$25.00
−60%
GPT-5.6 Sol
gpt-5.6-sol
complex reasoning, agentic coding, long-horizon tasks
ReasoningTools
1M
51
$1.60$5.00
$8.00$30.00
−20%
hy4-preview
coding & agentic workflows, reasoning (preview)
ReasoningTools
1M
$0.6672$0.834
$2.0008$2.501
−20%
hy3
general-purpose coding, agentic workflows
ReasoningTools
262K
$0.091
$0.363
mimo-v2.6-flash
fast low-cost coding, agents, image input
ReasoningToolsVision
1M
$0.112$0.14
$0.224$0.28
−20%
mimo-v2.5
cost-efficient everyday coding, agentic workflows
ReasoningTools
1M
33
$0.112
$0.224
qwen3.6-plus
thinks-by-default flagship
ReasoningTools
1M
$0.26$0.325
$1.56$1.95
−20%
qwen3.7-plus
Qwen 3.7 agent flagship, long-horizon coding
ReasoningToolsVision
1M
$0.256$0.32
$1.024$1.28
−20%
qwen3.6-35b
long-context RAG, 35B MoE
ReasoningTools
262K
$0.112$0.14
$0.80$1.00
−20%
kimi-k2.7-code
Long-horizon agentic coding
ReasoningTools
256K
$0.584$0.73
$2.80$3.50
−20%
GPT-5.6 Terra
gpt-5.6-terra
everyday coding, reasoning, balanced agentic
ReasoningTools
1M
47
$1.60$1.00
$9.60$6.00
−20%
GPT-5.6 Luna
gpt-5.6-luna
high-volume chat, classification, lightweight agentic
ReasoningTools
1M
43
$0.16$0.10
$0.96$0.60
−20%
Grok 4.5
grok-4.5
coding, knowledge work, STEM
ReasoningTools
500K
45
$1.60$2.00
$4.80$6.00
−20%
grok-4.6
frontier coding, knowledge work, STEM, visual analysis
ReasoningToolsVision
500K
51
$2.00 ($4.00 >200K)
$6.00 ($12.00 >200K)
Claude Fable 5
claude-fable-5
deep reasoning, long-horizon agentic
ReasoningToolsVision
1M
53
$4.00$10.00
$20.00$50.00
−60%
Claude Opus 4.8
claude-opus-4-8
top-tier reasoning, coding, agentic
ReasoningToolsVision
1M
$2.00$5.00
$10.00$25.00
−60%
Claude Sonnet 5
claude-sonnet-5
newest Sonnet, stronger reasoning & coding, below list price
ReasoningToolsVision
1M
45
$1.00$3.00
$5.00$15.00
−67%
Claude Haiku 4.5
claude-haiku-4-5
fast, low-cost, high-volume tasks
ToolsVision
200K
$0.40$1.00
$2.00$5.00
−60%
gemini-3.7-flash
multimodal Flash for agents
ReasoningToolsVision
1M
45
$0.6375$0.75
$3.1875$3.75
−15%
gemini-3.6-flash
current general-purpose Flash GA
ReasoningToolsVision
1M
40
$0.6375$0.75
$3.1875$3.75
−15%
gemini-3.5-flash-lite
current low-cost, high-volume workloads
ToolsVision
1M
28
$0.255$0.30
$2.125$2.50
−15%
gemini-3.5-flash
next-gen Flash GA
ReasoningToolsVision
1M
$1.275$1.50
$7.65$9.00
−15%
gemini-3.1-pro-preview
flagship reasoning
ReasoningToolsVision
1M
37
$1.70$2.00
$10.20$12.00
−15%
jev-1.13
Typed decisions, not chat: routing, classification and scoring with calibrated probabilities via /v1/systemone
System One
32K
$0.042
Free
flux.2-pro
flagship image generation
Image
—
—
$0.027/img$0.031/img
flux.1-schnell
fast, high-volume image drafts
Image
—
—
$0.003/img
sdxl-turbo
cheapest, fastest image previews
Image
—
—
$0.003/img
flux.2-klein
balanced open FLUX.2 image generation
Image
—
—
$0.02/img
Gemini 3 Pro Image
gemini-3-pro-image
GA pro-grade image generation
VisionImage
1M
$1.70$2.00
$10.20$12.00
per image $0.114/img$0.134/img
−15%

Prices are exact per-token rates, not rounded — some carry more decimals than others. Each struck-through figure is that row's reference list price and links to the page that publishes it, so every line can be checked. "At list" means we sell at the reference price; "mixed" means we are below it on one side and above on the other. Every row checkable, every model verifiable →

Building an agent? Fund a key programmatically with USDC — no account, no card. x402 docs →

python
1# Two lines. That is the whole migration.
2from openai import OpenAI
3 
4client = OpenAI(
5 base_url="https://api.quicksilverpro.io/v1",
6 api_key="your-api-key",
7)
FAQ

Common questions

QuickSilver Pro is an OpenAI-compatible inference API. The current catalog is DeepSeek V4 Flash, DeepSeek V4.1 Flash, DeepSeek V4 Pro, Qwen3.8 Max, Qwen3.8 Max Prime, Qwen3.8 Omni Flash, Qwen3.7 Max, Qwen3.7 Plus, Qwen3.7 Flash, Qwen3.6 Plus, Qwen3.6-35B-A3B, Qwen3.8 27B, Qwen3.8 Flash Next, Kimi K2.6, Kimi K2.7 Code, Kimi K3, Muse Spark 1.3, Muse Spark 1.2, Muse Glimmer 30B, GLM 5.3, GLM 5.3 Prime, GLM 5.3 Flash, GLM 5.2, Nemotron 3.5 Lightning, GPT-OSS 120B, GPT-6 Astra, GPT-6 Luna, GPT-6.1 Sol, GPT-6 Sol, GPT-5.6 Luna, GPT-5.6 Terra, GPT-5.6 Sol, Grok 4.5, Grok 4.6, Grok 4.7, MiniMax M3, MiMo-V2.6-Pro, MiMo-V2.6-Flash, MiMo-V2.5, Hy4 Preview, Hy3, Jev 1.13, Claude Opus 5.5, Claude Opus 5, Claude Fable 5.1, Claude Fable 5, Claude Opus 4.8, Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.1 Pro Preview, Gemini 3 Pro Image, Gemini 3 Flash Preview, Gemini 3.1 Flash Lite, FLUX.2 Pro, FLUX.1 Schnell, SDXL Turbo, FLUX.2 Klein, Qwen-Image Max, Seedream 5.0 Pro, Seedream 4 and Bria FIBO 1.5, served through one endpoint and one API key.

V4 Flash revision 0731 is the current agent build. It has 1M context and supports Chat Completions and Responses.

Up to 20% below the standard published per-token list rate on most of the catalog, and 50–67% below on Claude. DeepSeek V4 Flash: $0.086 / $0.173. DeepSeek V4 Pro: $0.70 / $2.10. Kimi K3: $2.55 / $12.75. GLM 5.2: $1.12 / $3.52. Claude Opus 5: $2.00 / $10.00. GPT-5.6 Sol: $1.60 / $8.00. Grok 4.5: $1.60 / $4.80. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-5.6, Gemini and Grok included.

Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, and usage.cost accounting all work out of the box. json_schema strict mode is model-dependent: the Claude models do not support it, so a schema there is advisory and the enforced path is a tool with strict: true.

Yes, two ways. Every account gets a free trial balance to use in the browser chat at /chat — no card, no API key. And your first credit purchase is matched 100%, up to $50 in bonus credits. The bonus is added to your balance the next time you top up (any amount from $5): pay $20, then top up $5 later, and $20 of bonus arrives with it. A larger first payment still receives the $50 maximum. One per customer and payment card; standard pay-as-you-go after that.

Get your API key

Create an account, get your API key in 30 seconds.

Get API Key

Model launches & updates. A few emails a month.