Gemini 3 Flash Preview QuickSilver Pro पर
Gemini 3 Flash Preview केवल मौजूदा integrations के साथ compatibility के लिए उपलब्ध है। Production और नए workloads को gemini-3.6-flash पर migrate करें, जो Google का मौजूदा GA Flash मॉडल है।
एक नज़र में
Gemini 3.6 Flash पर migrate हो रहे integrations के लिए अस्थायी compatibility।
Pricing तुलना ($/1M tokens)
| Provider | Input | Output | QSP की तुलना में |
|---|---|---|---|
| QuickSilver Pro | $0.425 | $2.55 | सबसे कम कीमत |
| OpenRouter (google/gemini-3-flash-preview) | $0.50 | $3.00 | 15% कम |
| OpenAI (GPT-4o-mini) | $0.15 | $0.60 | 325% महँगा |
कब इस्तेमाल करें
3 Flash Preview तब चुनें जब Flash Lite पर्याप्त स्मार्ट न हो लेकिन 3.1 Pro ज़रूरत से ज़्यादा हो: agentic loop में कठिन coding turns, multi-step analysis, और long-context summarization जिसमें documents के बीच non-trivial reasoning हो। Output price अब भी 3.1 Pro Preview से ~4× कम है।
कब कोई और model चुनें
Revenue-critical production paths के लिए, जब तक 3 Flash promote नहीं होता, किसी Gemini GA मॉडल (3.5 Flash) को प्राथमिकता दें — preview semantics बदल सकती है। Low-cost high-volume chat के लिए $0.2125/$1.275 पर 3.1 Flash Lite काफ़ी सस्ता है। Top-tier reasoning के लिए 3.1 Pro Preview ($1.70/$10.20) या DeepSeek V4 Pro ($0.70/$2.10) पर जाएँ।
Quickstart (curl)
curl https://api.quicksilverpro.io/v1/chat/completions \
-H "Authorization: Bearer $QSP_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3-flash-preview",
"messages": [{"role": "user", "content": "Hello!"}]
}'OpenAI-संगत। base_url बदलकर एक लाइन में migration।
FAQ
Preview का मतलब है कि Google बिना पूर्व सूचना के output formats, thinking behavior या pricing बदल सकता है। Prototyping और A/B evals के लिए आज ही ship करें। जिन paths में behavior बदलने से customers का काम टूट सकता है, वहाँ इसे feature flag के पीछे रखें और जब तक Google 3 Flash को GA label नहीं देता, fallback के रूप में 3.5 Flash GA पर pin करें।
3 Flash Preview निचला, preview-tier Flash है (प्रति 1M tokens $0.425/$2.55); 3.5 Flash उससे ऊपर का GA कदम है ($1.275/$7.65), ज़्यादा मज़बूत reasoning और committed-stable behavior के साथ। Cost-sensitive prototyping और evals के लिए 3 Flash Preview इस्तेमाल करें; जब production stability चाहिए तब 3.5 Flash GA पर जाएँ। तय करने से पहले अपने traffic पर side-by-side eval चलाएँ।
QuickSilver Pro पर Gemini 3 Flash Preview की कीमत प्रति 1M tokens $0.425 input / $2.55 output है — Vertex retail और OpenRouter की कीमत से ~15% कम; उनकी कीमत है $0.50/$3.00। OpenAI-compatible API; किसी दूसरे provider से switch करने के लिए OpenAI SDK में बस base_url + key बदलें।