होम/Models/Jev 1.13
System One32K context

Jev 1.13 QuickSilver Pro पर

Jev 1.13 TypeSafe का System One मॉडल है। यह text नहीं लिखता: आप एक state और typed सवालों का एक set भेजते हैं, और यह हर सवाल का जवाब एक ही forward pass में देता है — आपके options में से एक चुनना, state को आपके scale पर रखना, या हाँ की ओर झुकी probability लौटाना — साथ में calibrated probabilities के साथ। इसे POST /v1/systemone पर उसी QuickSilver Pro key और USD balance से call करें जो बाकी catalog के लिए है: $0.042 प्रति million input tokens, और output tokens मुफ़्त हैं।

प्रति 1M input tokens $0.042 · output मुफ़्त
लेखक:Raullen Chai·अपडेट:

एक नज़र में

Context
32K tokens
Input / 1M
$0.042
Output / 1M
मुफ़्त
Endpoint
/v1/systemone

Routing, classification, triage और guardrail के फ़ैसले calibrated probabilities के रूप में — parse करने को कुछ नहीं।

Pricing तुलना ($/1M tokens)

ProviderInputOutputQSP की तुलना में
QuickSilver Pro$0.042मुफ़्त—
TypeSafe list price (jev-1.13)$0.042मुफ़्तबराबर

कब इस्तेमाल करें

Jev 1.13 वहाँ चुनें जहाँ आपके code को गद्य नहीं, फ़ैसला चाहिए: किसी request को सही मॉडल या queue पर route करना, tickets को classify या triage करना, किसी rubric के आधार पर content को score करना, moderation और guardrail checks, या किसी agent का अगला कदम चुनना। हर जवाब typed रूप में लौटता है — हर option की probability और एक confidence के साथ choice, एक score, या एक probability — इसलिए न parse करने के लिए कोई free text है और न repair करने के लिए कोई JSON, और आप किसी अकेले label पर भरोसा करने के बजाय probability पर threshold लगा सकते हैं। एक ही call में अधिकतम 64 सवाल एक state साझा कर सकते हैं, और चूँकि केवल input का bill बनता है, 1,000-token state पर एक फ़ैसले की लागत लगभग इतनी आती है: $0.00004।

कब कोई और model चुनें

Jev 1.13 text generate नहीं करता — न जवाब, न summaries, न code, न tool calls — और इस तक Chat Completions के ज़रिए नहीं पहुँचा जा सकता; किसी भी generative काम के लिए chat मॉडल इस्तेमाल करें। इसकी window 32K tokens की है, इसलिए लंबे documents के बारे में पूछने से पहले उन्हें छोटा करें या उनका सारांश बनाएँ। Responses stream नहीं होते और एक call में शुरू से अंत तक आम तौर पर लगभग 2–3 सेकंड लगते हैं, इसलिए tight loops में कई calls करने के बजाय कई सवाल एक ही request में रखें।

Quickstart: POST /v1/systemone (curl)

curl https://api.quicksilverpro.io/v1/systemone \
  -H "Authorization: Bearer $QSP_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "jev-1.13",
    "state": "hi",
    "questions": {
      "greeting": {
        "type": "choice",
        "instructions": "Is the state a greeting?",
        "criteria": {"yes": "it is a greeting", "no": "it is not a greeting"}
      }
    }
  }'

# {"id":"so-45a870bb532d4a72b41404cc8911c72e","model":"jev-1.13",
#  "answers":{"greeting":{"type":"choice","choice":"yes",
#    "probabilities":{"no":0.07,"yes":0.93},"confidence":0.87}},
#  "usage":{"input_tokens":314,"output_tokens":33,"cost":...,"currency":"USD"}}
python (requests)
import os, requests

resp = requests.post(
    "https://api.quicksilverpro.io/v1/systemone",
    headers={"Authorization": f"Bearer {os.environ['QSP_KEY']}"},
    json={
        "model": "jev-1.13",
        "state": {"ticket": "I was charged twice for my subscription this month."},
        "questions": {
            "queue": {
                "type": "choice",
                "instructions": "Which team should handle this ticket?",
                "criteria": {
                    "billing": "payments, charges, refunds, invoices",
                    "technical": "bugs, errors, integrations",
                    "account": "login, email, profile changes",
                },
            },
            "urgent": {
                "type": "noul",
                "instructions": "Does this ticket need a reply within the hour?",
            },
        },
    },
    timeout=30,
)
resp.raise_for_status()
data = resp.json()
queue = data["answers"]["queue"]
print(queue["choice"], queue["probabilities"])  # picked queue + per-option probabilities
print(data["answers"]["urgent"]["noul"])        # probability-like number
print(data["usage"])                              # input_tokens, output_tokens, cost

यह System One API है, Chat Completions नहीं: state और typed questions भेजें, probabilities के साथ typed answers पाएँ। बाकी सभी models वाली ही key और balance। System One के पूरे docs →

FAQ

ऐसा मॉडल जो text लिखने के बजाय फ़ैसले करता है। आप एक state भेजते हैं — कोई भी JSON, string या object — और साथ में typed सवालों का एक set, और Jev 1.13 हर सवाल का जवाब एक ही forward pass में calibrated probabilities के साथ देता है। सवाल तीन तरह के होते हैं: choice (आपके बताए options में से एक चुनना, हर option की probability के साथ), score (state को आपके बताए scale पर रखना) और noul (हाँ की ओर झुके सवाल के लिए probability जैसी एक संख्या)।

POST https://api.quicksilverpro.io/v1/systemone करें, अपनी QuickSilver Pro key को Bearer token के रूप में और model, state व questions वाली JSON body के साथ — प्रति call अधिकतम 64 सवाल। यह एक सादा HTTPS JSON endpoint है, इसलिए curl, requests या httpx सब काम करते हैं। यह OpenAI Chat Completions surface का हिस्सा नहीं है: jev-1.13 पर chat call reject हो जाती है, और streaming supported नहीं है। /docs/systemone guide में पूरा schema, तीनों तरह के सवाल और error codes दिए गए हैं।

केवल input tokens पर, $0.042 प्रति million की दर से — output tokens मुफ़्त हैं। हर response का usage block input_tokens, output_tokens और उस call की USD cost बताता है, जो उसी prepaid balance से कटती है जिससे बाकी सभी models की। यह TypeSafe की अपनी list rate है; QuickSilver Pro कोई markup नहीं जोड़ता।

जवाब आपके चुने हुए question IDs से keyed होते हैं। choice सवाल चुना गया option, हर option की probability और एक confidence लौटाता है — उदाहरण के लिए {"type":"choice","choice":"yes","probabilities":{"no":0.07,"yes":0.93},"confidence":0.87}। noul सवाल एक ही संख्या लौटाता है, जैसे {"type":"noul","noul":0.69}। Response में एक id, model और call की cost वाला usage block भी होता है।

Jev 1.13 को double credits के साथ आज़माएँ — $50 तक bonus credits

API Key पाएँ