Grok 4.7 QuickSilver Pro पर
Grok 4.7 coding, agentic tasks, और knowledge work के लिए xAI का flagship है, Grok 4.6 का उत्तराधिकारी। यह लंबे समय तक चलने वाली software engineering के लिए बना है और जवाब देने से पहले अपना काम खुद जाँचता है। यह 500K-token context में text और images लेता है और streaming, tools, और structured output को support करता है। QuickSilver Pro $2.00 input / $6.00 output प्रति million tokens लेता है — xAI की प्रकाशित list price, बिना किसी QSP markup के।
एक नज़र में
लंबे समय तक चलने वाली agentic coding और knowledge work — self-verification और 500K-token context के साथ।
Pricing तुलना ($/1M tokens)
| Provider | Input | Output | QSP की तुलना में |
|---|---|---|---|
| QuickSilver Pro | $2.00 | $6.00 | — |
| xAI list price (Grok API) | $2.00 | $6.00 | बराबर |
कब इस्तेमाल करें
Grok 4.7 को लंबे समय तक चलने वाली software engineering, multi-step agent loops, technical research, और ऐसे knowledge work के लिए इस्तेमाल करें जहाँ मॉडल को अपना output लौटाने से पहले खुद verify करना चाहिए। यह उसी OpenAI-compatible QSP endpoint के ज़रिए Chat Completions और Responses दोनों को support करता है, जिसमें streaming, function calling, JSON Schema output, और image input शामिल हैं।
कब कोई और model चुनें
Reasoning हमेशा चालू रहती है और बंद नहीं की जा सकती — reasoning.enabled=false reject हो जाता है; गहराई के बदले latency और लागत घटाने के लिए reasoning_effort इस्तेमाल करें। Routine chat, extraction, या high-volume automation के लिए Flash-tier मॉडल आम तौर पर सस्ता पड़ेगा। 200,000 tokens से ऊपर के prompts के लिए पूरी request का bill $4 input और $12 output प्रति million tokens पर बनता है, इसलिए जब पूरे 500K context की ज़रूरत न हो तो retrieval या prompt compaction इस्तेमाल करें।
Quickstart (curl)
curl https://api.quicksilverpro.io/v1/chat/completions \
-H "Authorization: Bearer $QSP_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "grok-4.7",
"messages": [{"role": "user", "content": "Hello!"}]
}'OpenAI-संगत। base_url बदलकर एक लाइन में migration।
FAQ
200,000 prompt tokens तक की requests के लिए इसकी कीमत $2.00 प्रति million input tokens, $6.00 प्रति million output tokens, और $0.50 प्रति million cached-input tokens है। 200,000 prompt tokens से ऊपर पूरी request का bill $4 input, $12 output, और $1 cached input प्रति million tokens पर बनता है। QuickSilver Pro बिना markup के xAI की प्रकाशित rates के बराबर रखता है।
नहीं। Grok 4.7 हमेशा reasoning करता है, और reasoning.enabled=false वाली request reject हो जाती है। यह कितनी गहराई से सोचे, इसे नियंत्रित करने के लिए reasoning_effort सेट करें — कम effort तेज़ होता है और कम output tokens इस्तेमाल करता है।
base_url=https://api.quicksilverpro.io/v1 सेट करें, अपनी QSP API key इस्तेमाल करें, और model="grok-4.7" सेट करें। `/v1/chat/completions` और `/v1/responses` दोनों supported हैं।
Grok 4.7, xAI की ओर से Grok 4.6 का उत्तराधिकारी है, जो लंबे समय तक चलने वाली software engineering और अपना काम खुद verify करने के लिए tune किया गया है, उसी कीमत पर: $2.00 / $6.00 प्रति million tokens। दोनों में image input, tools, और structured output के साथ 500K-token context है। Grok 4.6 उपलब्ध बना हुआ है, इसलिए switch करना एक लाइन का model बदलाव है।