यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्या है
NeMo is an open-source toolkit from NVIDIA for building, training and customising large models across language, speech and multimodal directions. It works closely with NVIDIA GPUs and acceleration libraries, providing training and inference components. It addresses the difficulty of assembling flows and configurations for large-scale training and fine-tuning on GPU clusters.
यह क्यों महत्वपूर्ण है
It ties the training and customisation flow to NVIDIA’s hardware and software stack, an entry point for that compute on the model-engineering side.
मुख्य विशिष्टताएँ
- Directions
- Language, speech, multimodal
- Use
- Training, fine-tuning and inference
संबंधित अवधारणाएँ
प्रशिक्षण एवं अनुमान अधोसंरचना
स्मृति तय करती है कि कितना बड़ा मॉडल प्रशिक्षित हो सकता है, और संचार तय करता है कि कितना समय लगेगा
अनुमान अनुकूलन एवं सर्विंग
प्रशिक्षण एक बार होता है, अनुमान दिन में अरबों बार — और पहला टोकन एवं थ्रूपुट प्रायः एक-दूसरे से टकराते हैं