यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्या है
Phi is a family of small language models introduced by Microsoft Research in June 2023. Its emphasis is data quality rather than model size: training on filtered “textbook-quality” text lets a few-billion-parameter model approach much larger ones on reasoning and coding. Phi-3 offers 3.8B, 7B and 14B sizes, and the smallest, Phi-3-mini, can run on-device. The models are released as open weights and are often used to probe the capability limits of small models.
यह क्यों महत्वपूर्ण है
It showed that curating and orchestrating data can substitute for parameter count to a surprising degree: a few-billion-parameter model trained on selected data can approach models an order of magnitude larger on some reasoning tasks.
मुख्य विशिष्टताएँ
- Parameters
- 3.8B / 7B / 14B (Phi-3 family)
- Training data
- Mainly filtered “textbook-quality” text
- Open weights
- Yes
- Released
- 2023-06
संबंधित क्षमताएँ
संबंधित अवधारणाएँ
प्री-ट्रेनिंग और फाइन-ट्यूनिंग
पहले विशाल अलेबल पाठ से भाषा सीखना, फिर थोड़े डेटा से विशेषज्ञ बनना — आधुनिक एआई का सबसे डेटा-कुशल प्रतिमान
Transformer आर्किटेक्चर
शब्द-दर-शब्द relay की जगह वह कक्ष जहाँ सब एक साथ बोलते हैं, जिससे दूर की निर्भरताएँ एक कदम पर आ जाती हैं
मॉडल संपीड़न
सटीकता लगभग बनाए रखते हुए मॉडल को छोटा, तेज़ और सस्ता बनाना — पर तीनों एक साथ कम ही मिलते हैं
समान उत्पाद
Llama
2023खुले भार के रास्ते को मुख्यधारा बनाने वाला मॉडल परिवार
Mistral Large
2024यूरोपीय ओपन-वेट लैब का फ़्लैगशिप वाणिज्यिक मॉडल
Qwen
2023अनेक आकारों और बहुविध संस्करणों वाला ओपन-वेट परिवार