Gemini 3.8 Flash QuickSilver Pro पर
Gemini 3.8 Flash production agents, coding, multimodal analysis और long-context काम के लिए Google का सबसे सक्षम Flash मॉडल है। High reasoning के साथ Artificial Analysis Intelligence Index पर इसका score 59 है, और यह 1M-token context में text और images स्वीकार करता है, streaming, tools और structured output के साथ। 31 दिसंबर 2026 तक QuickSilver Pro प्रति million tokens $0.6375 input / $3.1875 output चार्ज करता है — Google की promotional API price से 15% कम।
एक नज़र में
Agents और coding के लिए सबसे स्मार्ट Flash — 1M context, Google की 2026 promotional API price से 15% कम पर।
Pricing तुलना ($/1M tokens)
| Provider | Input | Output | QSP की तुलना में |
|---|---|---|---|
| QuickSilver Pro | $0.6375 | $3.1875 | सबसे कम कीमत |
| Google की promotional price (Gemini API promotional rate) | $0.75 | $3.75 | 15% कम |
| OpenAI (GPT-4o) | $2.50 | $10.00 | 68% कम |
कब इस्तेमाल करें
Gemini 3.8 Flash का उपयोग production agents, coding assistants, image understanding, document analysis, tool-calling workflows और उन लंबे prompts के लिए करें जिन्हें 1M-token context से फ़ायदा होता है। यह उसी OpenAI-compatible QSP endpoint के ज़रिए Chat Completions और Responses दोनों को support करता है।
कब कोई और model चुनें
साधारण classification या extraction के लिए, जहाँ reasoning quality से ज़्यादा लागत मायने रखती है, Gemini 3.5 Flash-Lite सस्ता है। कम token price पर text-only premium reasoning के लिए DeepSeek V4 Pro को परखें। Gemini 3.8 Flash की मौजूदा कीमत 31 दिसंबर 2026 तक promotional है, इसलिए budget-sensitive production workloads को 2027 से पहले pricing की समीक्षा कर लेनी चाहिए।
Quickstart (curl)
curl https://api.quicksilverpro.io/v1/chat/completions \
-H "Authorization: Bearer $QSP_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.8-flash",
"messages": [{"role": "user", "content": "Hello!"}]
}'OpenAI-संगत। base_url बदलकर एक लाइन में migration।
FAQ
31 दिसंबर 2026 तक इसकी कीमत प्रति million input tokens $0.6375 और प्रति million output tokens $3.1875 है। Cached input प्रति million tokens $0.06375 है। ये rates Google की promotional API prices से 15% कम हैं; promotion खत्म होने से पहले pricing की समीक्षा की जाएगी।
High reasoning के साथ Artificial Analysis Intelligence Index पर इसका score 59 है — Gemini 3.7 Flash से 3 points ऊपर और intelligence-vs-cost Pareto frontier पर — और agentic tool-use तथा coding evaluations में सुधार खास तौर पर मज़बूत है।
हाँ। यह text और image input स्वीकार करता है और streaming, reasoning, function calling और JSON Schema structured output को support करता है। इसका अधिकतम context 1M tokens है और अधिकतम output 65,536 tokens।
base_url=https://api.quicksilverpro.io/v1 सेट करें, अपनी QSP API key इस्तेमाल करें, और model="gemini-3.8-flash" सेट करें। `/v1/chat/completions` और `/v1/responses` दोनों supported हैं, streaming और tool calls सहित।