Home/Models/Claude Sonnet 5.5
1M contextMultimodalReasoningResponses API

Claude Sonnet 5.5 on QuickSilver Pro

Claude Sonnet 5.5: $1.00 input / $5.00 output per 1M tokens, 50% below Anthropic list ($2.00 / $10.00). Context: 1M tokens. Maximum output: 128K tokens. Available through Chat Completions and Responses. Capabilities: text input, text output, image input, streaming, reasoning and tool calling.

$1.00 input · $5.00 output per 1M tokens
Get API Key

Unlocks with your first top-up from $5

ByRaullen Chai·Updated

At a glance

Context
1M tokens
Input / 1M
$1.00
Output / 1M
$5.00
Thinks by default
Yes

For developers building coding assistants and tool-based workflows.

Pricing comparison ($/1M tokens)

ProviderInputOutputvs QSP
QuickSilver Pro$1.00$5.00lowest-cost
Anthropic list price (Anthropic API)$2.00$10.0050% lower

When to use

Use Claude Sonnet 5.5 for Sonnet reasoning, image input and tool calling. Context: 1M tokens; maximum output: 128K tokens. The gateway drops `temperature` and `top_p` and rewrites assistant prefill as a user continuation.

When to use something else

For base output savings, use Claude Haiku 5.5: 1M context / 128K maximum output tokens at 0.05× this model’s output price. Choose Claude Sonnet 5 for native forced `tool_choice`; its base output costs 1× this model. Above 200K prompt tokens, Claude Haiku 5.5 bills the whole request at $0.25 input / $1.25 output per million tokens.

Quickstart (curl)

curl https://api.quicksilverpro.io/v1/chat/completions \
  -H "Authorization: Bearer $QSP_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-sonnet-5-5",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
python (OpenAI SDK)
import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.quicksilverpro.io/v1",
    api_key=os.environ["QSP_KEY"],
)
response = client.chat.completions.create(
    model="claude-sonnet-5-5",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

OpenAI-compatible. One-line migration via base_url.

FAQ

Claude Sonnet 5.5 costs $1.00 input / $5.00 output per 1M tokens. Anthropic list is $2.00 / $10.00; the discount is 50% on both base rates.

The catalog has no separate long-context tier for Claude Sonnet 5.5. Base rates apply within its 1M-token context.

The gateway drops `temperature` and `top_p`. This model rejects a trailing assistant prefill, so the gateway rewrites it as a user continuation turn carrying the prefix.

A forced `tool_choice` becomes `auto` plus an instruction to call the tool; a call is steered, not guaranteed. Claude `response_format` with `json_schema` is advisory. Use a tool with `strict: true` for enforced argument shape, and handle replies without a tool call.

Catalog capabilities: text input, text output, image input, streaming, reasoning and tool calling. For Claude image requests, the gateway fetches public remote image URLs and inlines the image data. An inaccessible or rejected URL fails the request; inline base64 is also supported.

Set `base_url="https://api.quicksilverpro.io/v1"`, use your QSP API key, and set `model="claude-sonnet-5-5"`. Supported APIs: Chat Completions and Responses. See the runnable quickstart on this page and the API docs.

Claude · Models

Claude Sonnet 5.5

Unlocks with your first top-up from $5

Get API Key