Câu 20: NCP-AIO: NCP - AI Operations
You are deploying an AI workload on a Kubernetes cluster that requires access to GPUs for training deep learning models. However, the pods are not able to detect the GPUs on the nodes. What would be the first step to troubleshoot this issue?
Nội dung câu hỏi
You are deploying an AI workload on a Kubernetes cluster that requires access to GPUs for training deep learning models. However, the pods are not able to detect the GPUs on the nodes. What would be the first step to troubleshoot this issue?
Các lựa chọn
Đáp án được giữ gọn theo nhãn A, B, C, D trong phần bình chọn tương tác.
- A. Verify that the NVIDIA GPU Operator is installed and running on the cluster. — đáp án hiện tại
- B. Ensure that all pods are using the latest version of TensorFlow or PyTorch.
- C. Check if the nodes have sufficient memory allocated for AI workloads.
- D. Increase the number of CPU cores allocated to each pod to ensure better resource utilization.
Cộng đồng
0 bình luận công khai. Tên thành viên được ẩn một phần.
Chưa có bình luận. Mở giao diện tương tác để bắt đầu thảo luận.