Toàn văn của trang này được trình bày bằng tiếng Anh; tiêu đề và phần dẫn đã được bản địa hóa.
NÓ LÀ GÌ
NeMo is an open-source toolkit from NVIDIA for building, training and customising large models across language, speech and multimodal directions. It works closely with NVIDIA GPUs and acceleration libraries, providing training and inference components. It addresses the difficulty of assembling flows and configurations for large-scale training and fine-tuning on GPU clusters.
Vì sao đáng ghi nhớ
It ties the training and customisation flow to NVIDIA’s hardware and software stack, an entry point for that compute on the model-engineering side.
Thông số chính
- Directions
- Language, speech, multimodal
- Use
- Training, fine-tuning and inference
Khái niệm liên quan
Hạ tầng huấn luyện và suy luận
Bộ nhớ quyết định mô hình lớn đến đâu, còn truyền thông quyết định mất bao lâu
Tối ưu suy luận và phục vụ
Huấn luyện chỉ một lần, suy luận diễn ra hàng tỷ lần mỗi ngày — và token đầu tiên lẫn thông lượng thường kéo ngược nhau