Le texte intégral est présenté en anglais ; le titre et le résumé sont localisés.
CE QUE C'EST
NeMo is an open-source toolkit from NVIDIA for building, training and customising large models across language, speech and multimodal directions. It works closely with NVIDIA GPUs and acceleration libraries, providing training and inference components. It addresses the difficulty of assembling flows and configurations for large-scale training and fine-tuning on GPU clusters.
Pourquoi il compte
It ties the training and customisation flow to NVIDIA’s hardware and software stack, an entry point for that compute on the model-engineering side.
Caractéristiques clés
- Directions
- Language, speech, multimodal
- Use
- Training, fine-tuning and inference
Concepts liés
Infrastructure d’entraînement et d’inférence
La mémoire décide de la taille du modèle entraînable, la communication de la durée — le calcul brut est rarement le goulot
Optimisation et service d’inférence
L’entraînement a lieu une fois, l’inférence des milliards de fois par jour — et le premier token s’oppose souvent au débit