DeepInfra → QuickSilver Pro
DeepInfra is already the budget option among DeepSeek resellers — QuickSilver Pro serves the latest V4 wave (V4 Flash for low-cost chat, V4 Pro for reasoning) with per-token cache pricing. Both APIs are OpenAI-compatible, so the move is a base-URL swap. For the full side-by-side analysis, see /vs/deepinfra.
The steps
- 1
Get a QuickSilver Pro API key
Sign up at quicksilverpro.io/dashboard. First top-up bonus: we match your first top-up 100%, up to $50 — added to your balance when you top up again (any amount from $5).
- 2
Change the base URL
In your OpenAI SDK init, swap the base_url. Note DeepInfra's OpenAI-compatible path ends in /v1/openai.
- base_url="https://api.deepinfra.com/v1/openai" + base_url="https://api.quicksilverpro.io/v1" - 3
Swap the API key
Replace your DeepInfra token with a QuickSilver Pro key.
- api_key=os.environ["DEEPINFRA_TOKEN"], + api_key=os.environ["QSP_KEY"], - 4
Rename model IDs
DeepInfra prefixes model IDs with the originating org. Drop the prefix and use the QuickSilver Pro short name.
DeepInfra QuickSilver Pro deepseek-ai/DeepSeek-V4-Flash deepseek-v4-flash deepseek-ai/DeepSeek-V4-Pro deepseek-v4-pro - 5
Test your core flows end-to-end
Run one representative request for each feature you use — chat, streaming, tool / function calling, and json_schema strict mode. Apart from the known difference below, any behavioral difference is a bug — report it.
One known difference: json_schema strict mode is model-dependent. The Claude models do not support it, so a schema there is advisory and the enforced path is a tool with
strict: true— structured output.
Full before/after
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.deepinfra.com/v1/openai",
api_key=os.environ["DEEPINFRA_TOKEN"],
)
r = client.chat.completions.create(
model="deepseek-ai/DeepSeek-V4-Flash",
messages=[{"role": "user", "content": "Hi"}],
)import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.quicksilverpro.io/v1",
api_key=os.environ["QSP_KEY"],
)
r = client.chat.completions.create(
model="deepseek-v4-flash",
messages=[{"role": "user", "content": "Hi"}],
)What you'll pay after switching
Per 1M tokens, input / output. QuickSilver Pro rates vs DeepInfra's published per-token pricing.
| Model | QuickSilver Pro | DeepInfra | Savings |
|---|---|---|---|
| DeepSeek V4 Pro | $0.70 / $2.10 | $1.30 / $2.60 | ~19% output |
Common migration pitfalls
Migrating from DeepInfra — FAQ
Other migration guides
Need help?
Email [email protected] — a human replies usually within 4 hours. For the broader analysis, see QuickSilver Pro vs DeepInfra.
Start saving in 5 minutes
First top-up matched 100%, up to $50 in bonus credits, added on your next top-up. Keep your code on the OpenAI SDK — only the base URL and key change.
Get API Key