O texto completo é apresentado em inglês; o título e o resumo estão traduzidos.
O QUE É
NeMo is an open-source toolkit from NVIDIA for building, training and customising large models across language, speech and multimodal directions. It works closely with NVIDIA GPUs and acceleration libraries, providing training and inference components. It addresses the difficulty of assembling flows and configurations for large-scale training and fine-tuning on GPU clusters.
Por que vale a pena lembrar
It ties the training and customisation flow to NVIDIA’s hardware and software stack, an entry point for that compute on the model-engineering side.
Especificações-chave
- Directions
- Language, speech, multimodal
- Use
- Training, fine-tuning and inference
Conceitos relacionados
Infraestrutura de treinamento e inferência
A memória decide o tamanho do modelo treinável; a comunicação, quanto tempo leva
Otimização e serviço de inferência
O treino ocorre uma vez; a inferência, bilhões de vezes por dia — e o primeiro token e a vazão costumam se opor