होम/Models/Gemini 3.7 Flash
1M contextMultimodalReasoningResponses API

Gemini 3.7 Flash QuickSilver Pro पर

Gemini 3.7 Flash production agents, coding, multimodal analysis और long-context काम के लिए Google का सबसे नया Flash मॉडल है। यह 1M-token context में text और images स्वीकार करता है और reasoning, streaming, tools और structured output को support करता है। 31 दिसंबर 2026 तक QuickSilver Pro प्रति million tokens $0.6375 input / $3.1875 output चार्ज करता है — Google की promotional API price से 15% कम।

प्रति 1M tokens: input $0.6375 · output $3.1875
लेखक:Raullen Chai·अपडेट:

एक नज़र में

Context
1M tokens
Input / 1M
$0.6375
Output / 1M
$3.1875
Default में सोचता है
हाँ

1M context के साथ तेज़ multimodal agents और coding, Google की 2026 promotional API price से 15% कम पर।

Pricing तुलना ($/1M tokens)

ProviderInputOutputQSP की तुलना में
QuickSilver Pro$0.6375$3.1875सबसे कम कीमत
Google की promotional price (Gemini API promotional rate)$0.75$3.7515% कम
OpenAI (GPT-4o)$2.50$10.0068% कम

कब इस्तेमाल करें

Gemini 3.7 Flash का उपयोग production agents, coding assistants, image understanding, document analysis, tool-calling workflows और उन लंबे prompts के लिए करें जिन्हें 1M-token context से फ़ायदा होता है। यह उसी OpenAI-compatible QSP endpoint के ज़रिए Chat Completions और Responses दोनों को support करता है।

कब कोई और model चुनें

साधारण classification या extraction के लिए, जहाँ reasoning quality से ज़्यादा लागत मायने रखती है, Gemini 3.5 Flash-Lite सस्ता है। कम token price पर text-only premium reasoning के लिए DeepSeek V4 Pro को परखें। Gemini 3.7 Flash की मौजूदा कीमत 31 दिसंबर 2026 तक promotional है, इसलिए budget-sensitive production workloads को 2027 से पहले pricing की समीक्षा कर लेनी चाहिए।

Quickstart (curl)

curl https://api.quicksilverpro.io/v1/chat/completions \
  -H "Authorization: Bearer $QSP_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3.7-flash",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

OpenAI-संगत। base_url बदलकर एक लाइन में migration।

FAQ

31 दिसंबर 2026 तक इसकी कीमत प्रति million input tokens $0.6375 और प्रति million output tokens $3.1875 है। Cached input प्रति million tokens $0.06375 है। ये rates Google की promotional API prices से 15% कम हैं; promotion खत्म होने से पहले pricing की समीक्षा की जाएगी।

हाँ। यह text और image input स्वीकार करता है और streaming, reasoning, function calling और JSON Schema structured output को support करता है। इसका अधिकतम context 1M tokens है और अधिकतम output 65,536 tokens।

base_url=https://api.quicksilverpro.io/v1 सेट करें, अपनी QSP API key इस्तेमाल करें, और model="gemini-3.7-flash" सेट करें। `/v1/chat/completions` और `/v1/responses` दोनों supported हैं, streaming और tool calls सहित।

हाँ। QuickSilver Pro मॉडल को लगातार monitor करता है, और launch के समय production endpoint पर non-streaming Chat Completions, streaming Chat Completions, Responses API, streaming Responses और forced function calling को verify किया गया था।

Gemini 3.7 Flash को double credits के साथ आज़माएँ — $50 तक bonus credits

API Key पाएँ