होम/Models/Grok 4.7
500K contextMultimodalReasoningResponses API

Grok 4.7 QuickSilver Pro पर

Grok 4.7 coding, agentic tasks, और knowledge work के लिए xAI का flagship है, Grok 4.6 का उत्तराधिकारी। यह लंबे समय तक चलने वाली software engineering के लिए बना है और जवाब देने से पहले अपना काम खुद जाँचता है। यह 500K-token context में text और images लेता है और streaming, tools, और structured output को support करता है। QuickSilver Pro $2.00 input / $6.00 output प्रति million tokens लेता है — xAI की प्रकाशित list price, बिना किसी QSP markup के।

प्रति 1M tokens: input $2.00 · output $6.00
लेखक:Raullen Chai·अपडेट:

एक नज़र में

Context
500K tokens
Input / 1M
$2.00
Output / 1M
$6.00
Default में सोचता है
हाँ

लंबे समय तक चलने वाली agentic coding और knowledge work — self-verification और 500K-token context के साथ।

Pricing तुलना ($/1M tokens)

ProviderInputOutputQSP की तुलना में
QuickSilver Pro$2.00$6.00—
xAI list price (Grok API)$2.00$6.00बराबर

कब इस्तेमाल करें

Grok 4.7 को लंबे समय तक चलने वाली software engineering, multi-step agent loops, technical research, और ऐसे knowledge work के लिए इस्तेमाल करें जहाँ मॉडल को अपना output लौटाने से पहले खुद verify करना चाहिए। यह उसी OpenAI-compatible QSP endpoint के ज़रिए Chat Completions और Responses दोनों को support करता है, जिसमें streaming, function calling, JSON Schema output, और image input शामिल हैं।

कब कोई और model चुनें

Reasoning हमेशा चालू रहती है और बंद नहीं की जा सकती — reasoning.enabled=false reject हो जाता है; गहराई के बदले latency और लागत घटाने के लिए reasoning_effort इस्तेमाल करें। Routine chat, extraction, या high-volume automation के लिए Flash-tier मॉडल आम तौर पर सस्ता पड़ेगा। 200,000 tokens से ऊपर के prompts के लिए पूरी request का bill $4 input और $12 output प्रति million tokens पर बनता है, इसलिए जब पूरे 500K context की ज़रूरत न हो तो retrieval या prompt compaction इस्तेमाल करें।

Quickstart (curl)

curl https://api.quicksilverpro.io/v1/chat/completions \
  -H "Authorization: Bearer $QSP_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "grok-4.7",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

OpenAI-संगत। base_url बदलकर एक लाइन में migration।

FAQ

200,000 prompt tokens तक की requests के लिए इसकी कीमत $2.00 प्रति million input tokens, $6.00 प्रति million output tokens, और $0.50 प्रति million cached-input tokens है। 200,000 prompt tokens से ऊपर पूरी request का bill $4 input, $12 output, और $1 cached input प्रति million tokens पर बनता है। QuickSilver Pro बिना markup के xAI की प्रकाशित rates के बराबर रखता है।

नहीं। Grok 4.7 हमेशा reasoning करता है, और reasoning.enabled=false वाली request reject हो जाती है। यह कितनी गहराई से सोचे, इसे नियंत्रित करने के लिए reasoning_effort सेट करें — कम effort तेज़ होता है और कम output tokens इस्तेमाल करता है।

base_url=https://api.quicksilverpro.io/v1 सेट करें, अपनी QSP API key इस्तेमाल करें, और model="grok-4.7" सेट करें। `/v1/chat/completions` और `/v1/responses` दोनों supported हैं।

Grok 4.7, xAI की ओर से Grok 4.6 का उत्तराधिकारी है, जो लंबे समय तक चलने वाली software engineering और अपना काम खुद verify करने के लिए tune किया गया है, उसी कीमत पर: $2.00 / $6.00 प्रति million tokens। दोनों में image input, tools, और structured output के साथ 500K-token context है। Grok 4.6 उपलब्ध बना हुआ है, इसलिए switch करना एक लाइन का model बदलाव है।

Grok 4.7 को double credits के साथ आज़माएँ — $50 तक bonus credits

API Key पाएँ