कोड जनरेशन
विवरण से चलने योग्य कोड लिखना
यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्षमता क्या है
Takes a description — write a function that reads a CSV into a list of dicts and skips blank lines — and outputs the corresponding source. Unlike code completion it targets a new feature or a whole piece of logic, usually written from scratch; unlike reasoning the deliverable is compilable, runnable code rather than a textual conclusion.
तकनीकी रूप से कैसे
The base is a language model pre-trained on large code corpora; strict syntax makes code easier to verify by execution than prose, so unit-test pass signals work well as reward for reinforcement learning or rejection-sampling fine-tuning. Training and evaluation use problems with test cases such as HumanEval and MBPP, and generation can sample several candidates and discard the ones that fail.
प्रतिनिधि उत्पाद
11GPT-4o
2024मूल रूप से बहुविध सामान्य मॉडल — पाठ, चित्र और ऑडियो एक ही द्वार से
Claude
2023लंबे संदर्भ और सुरक्षा-संरेखण के लिए जाना जाने वाला सामान्य संवाद मॉडल
DeepSeek-V3
2024671B पैरामीटर वाला ओपन-वेट MoE, प्रति टोकन केवल 37B सक्रिय
Qwen
2023अनेक आकारों और बहुविध संस्करणों वाला ओपन-वेट परिवार
Gemini
2023मूल रूप से बहुविध, अति-लंबे संदर्भ के लिए बना सामान्य मॉडल
GitHub Copilot
2021एडिटर में संदर्भ के अनुसार कोड पूरा करता और बदलता है
o3
2025उत्तर देने से पहले लंबी तर्क-श्रृंखला, अनुमान के समय गणना से अधिक सटीकता
Claude Code
2025टर्मिनल से बहु-चरणीय कोडिंग कार्य पूरे करने वाला एजेंट
DeepSeek-R1
2025तर्क-श्रृंखला पर RL से प्रशिक्षित तर्क मॉडल, भार MIT लाइसेंस के तहत खुले
Devin
2024अपने शेल, एडिटर और ब्राउज़र वाला स्वायत्त कोडिंग एजेंट
Cursor
2023डेस्कटॉप कोड एडिटर जो पूरे रिपॉज़िटरी को संदर्भ बनाता है
संबंधित संस्थान
सामान्य उपयोग
- New features and utility scripts
- Data cleaning and transformation scripts
- Test cases and project scaffolding
- Porting code across languages and frameworks
इसका मूल्यांकन कैसे होता है
- pass@k
- Share of problems with at least one sample passing all tests out of k
- Benchmark pass rate
- Pass rate on fixed sets such as HumanEval
- Compile and lint pass rate
- Whether generated code passes the compiler and type checks as-is
सीमाएँ और कठिनाइयाँ
- It invents libraries or methods, producing plausible interfaces that do not exist
- Code can pass unit tests yet leave holes in edge cases, error handling and security
- Multi-file changes are inconsistent, with call sites disagreeing with definitions
इसके पीछे की अवधारणाएँ
Transformer आर्किटेक्चर
शब्द-दर-शब्द relay की जगह वह कक्ष जहाँ सब एक साथ बोलते हैं, जिससे दूर की निर्भरताएँ एक कदम पर आ जाती हैं
प्री-ट्रेनिंग और फाइन-ट्यूनिंग
पहले विशाल अलेबल पाठ से भाषा सीखना, फिर थोड़े डेटा से विशेषज्ञ बनना — आधुनिक एआई का सबसे डेटा-कुशल प्रतिमान
प्रॉम्प्ट इंजीनियरिंग और अलाइनमेंट
मॉडल को उपयोगी, ईमानदार और हानिरहित बनाना उसे केवल बड़ा करने से कठिन है