Toàn văn của trang này được trình bày bằng tiếng Anh; tiêu đề và phần dẫn đã được bản địa hóa.
NÓ LÀ GÌ
CUDA is the parallel computing platform and programming model introduced by NVIDIA in 2007 that lets developers use GPUs for general-purpose computing. It provides language extensions, a compiler, and acceleration libraries such as cuDNN and cuBLAS. It addresses efficiently mapping highly parallel computation such as neural-network training and inference onto GPUs.
Vì sao đáng ghi nhớ
It has been the practical programming interface for AI compute for over a decade; the libraries and tools built around it form NVIDIA’s deepest ecosystem moat.
Thông số chính
- First release
- 2007
- Positioning
- GPU general-purpose parallel computing platform and programming model
- Companion libraries
- cuDNN, cuBLAS and others
Khái niệm liên quan
Hạ tầng huấn luyện và suy luận
Bộ nhớ quyết định mô hình lớn đến đâu, còn truyền thông quyết định mất bao lâu
Tối ưu suy luận và phục vụ
Huấn luyện chỉ một lần, suy luận diễn ra hàng tỷ lần mỗi ngày — và token đầu tiên lẫn thông lượng thường kéo ngược nhau