तर्क और विचार-शृंखला
कठिन समस्या को मध्यवर्ती चरणों में बाँटकर हल करना
यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्षमता क्या है
Takes a problem requiring several steps — a maths problem, a logic puzzle, a task needing a plan — and returns the answer together with intermediate steps. Unlike free generation the goal is correctness rather than fluency, and unlike ordinary conversation it spends much of its compute on an internal scratchpad, showing the user mainly the result and a short rationale.
तकनीकी रूप से कैसे
One family relies on prompting: examples or instructions ask the model to think before answering, writing steps out explicitly — chain-of-thought. Another relies on training: large-scale reinforcement learning on verifiable problems such as maths and code teaches the model to generate a longer deliberation before committing. Self-consistency voting over several sampled chains, and combining deliberation with tool calls or search, are now common too.
प्रतिनिधि उत्पाद
23o3
2025उत्तर देने से पहले लंबी तर्क-श्रृंखला, अनुमान के समय गणना से अधिक सटीकता
DeepSeek-R1
2025तर्क-श्रृंखला पर RL से प्रशिक्षित तर्क मॉडल, भार MIT लाइसेंस के तहत खुले
Gemini
2023मूल रूप से बहुविध, अति-लंबे संदर्भ के लिए बना सामान्य मॉडल
Claude
2023लंबे संदर्भ और सुरक्षा-संरेखण के लिए जाना जाने वाला सामान्य संवाद मॉडल
Qwen
2023अनेक आकारों और बहुविध संस्करणों वाला ओपन-वेट परिवार
MiniMax-M
2025हाइब्रिड अटेंशन और दस लाख टोकन संदर्भ वाला ओपन-वेट तर्क मॉडल
DeepSeek-V3
2024671B पैरामीटर वाला ओपन-वेट MoE, प्रति टोकन केवल 37B सक्रिय
GPT-4o
2024मूल रूप से बहुविध सामान्य मॉडल — पाठ, चित्र और ऑडियो एक ही द्वार से
Jamba
2024स्टेट-स्पेस मॉडल और Transformer को मिलाने वाला ओपन-वेट मॉडल
DBRX
2024Databricks का ओपन-वेट MoE भाषा मॉडल
Mistral Large
2024यूरोपीय ओपन-वेट लैब का फ़्लैगशिप वाणिज्यिक मॉडल
Gemini
2024खोज, ऑफ़िस और बहुविध मॉडल को एक चैट द्वार में समेटता है
Grok
2023सोशल प्लेटफ़ॉर्म के डेटा से जुड़ा संवाद मॉडल
Yi
2023चीनी-अंग्रेज़ी द्विभाषी ओपन-वेट मॉडल, अति-लंबे संदर्भ संस्करणों के साथ
Hunyuan
2023ओपन-वेट संस्करणों सहित टेनसेंट का सामान्य मॉडल परिवार
Doubao
2023बाइटडांस का सामान्य संवाद मॉडल और ऐप
Phi
2023चुने हुए डेटा से बने छोटे और कुशल मॉडल
Step
2023बहुविधता और ऑन-डिवाइस उपयोग पर लक्षित सामान्य मॉडल परिवार
Baichuan
2023चीनी-भाषा उपयोग के लिए ओपन-वेट सामान्य मॉडल
GLM
2023ऑटोरेग्रेसिव ब्लैंक-भरने वाले प्रीट्रेनिंग से शुरू हुआ चीनी सामान्य मॉडल
Llama
2023खुले भार के रास्ते को मुख्यधारा बनाने वाला मॉडल परिवार
ChatGPT
2022वह चैट विंडो जिसने बड़े भाषा मॉडल को हर किसी तक पहुँचाया
Pangu
2020हुवावे के Pangu फ़ाउंडेशन मॉडल परिवार
संबंधित संस्थान
सामान्य उपयोग
- Solving maths and physics problems
- Multi-step planning and scheduling
- Code and data problems needing derivation
- Reasoned choices under constraints
इसका मूल्यांकन कैसे होता है
- Math benchmark accuracy
- Correct-answer rate on sets such as GSM8K, MATH, AIME
- pass@k
- Share of problems solved by at least one of k samples
- Chain-of-thought faithfulness
- Whether the written steps actually reflect how the answer was reached
सीमाएँ और कठिनाइयाँ
- The written reasoning can be unfaithful: the answer comes first and a plausible rationale is added after
- In long chains an early misstep propagates, yet the final answer still sounds assured
- Deliberating at length on trivial questions inflates latency and cost
इसके पीछे की अवधारणाएँ
प्रॉम्प्ट इंजीनियरिंग और अलाइनमेंट
मॉडल को उपयोगी, ईमानदार और हानिरहित बनाना उसे केवल बड़ा करने से कठिन है
मानव प्रतिक्रिया से सुदृढ़ीकरण अधिगम
जब «अच्छा उत्तर» सूत्र में न लिखा जा सके, मनुष्य को पुरस्कार फलन बनने दें
Transformer आर्किटेक्चर
शब्द-दर-शब्द relay की जगह वह कक्ष जहाँ सब एक साथ बोलते हैं, जिससे दूर की निर्भरताएँ एक कदम पर आ जाती हैं