NVIDIA NCP-AII - NCP-AI Infrastructure Exam
Page: 2 / 34
Total 170 questions
Question #6 (Topic: Exam A)
An engineer needs to validate 400G DAC cable signal integrity in a DGX cluster.
Which CVT metric best identifies marginal cables needing replacement?
Which CVT metric best identifies marginal cables needing replacement?
A. Effective BER > 1.5E-254 during a ≤6-hour monitoring window.
B. Transceiver model matching QSFP-DD specifications.
C. Temperature fluctuations > 5°C during validation.
D. Lane power variance < 3dB across all transceivers.
Answer: A
Question #7 (Topic: Exam A)
A user encounters “permission denied” errors when running GPU-accelerated containers on a Secure Boot-enabled system.
What resolves this?
What resolves this?
A. Disable SELinux to relax unnecessary security policies.
B. Reinstall Docker without the NVIDIA runtime.
C. Enroll the MOK and sign NVIDIA kernel modules.
D. Run Docker with sudo for elevated privileges.
Answer: C
Question #8 (Topic: Exam A)
An InfiniBand administrator needs to run performance benchmarks on new devices added to the fabric.
What tool should be used to check the latency?
What tool should be used to check the latency?
A. tcpdump
B. ib_write_lat
C. perfmon
D. ibdiagnet
Answer: B
Question #9 (Topic: Exam A)
A cluster administrator is preparing to update the firmware on a DGX H100 system, including the GPU tray (baseboard).
What is the correct sequence of steps to perform a safe and successful firmware upgrade?
What is the correct sequence of steps to perform a safe and successful firmware upgrade?
A. Update the BMC and skip the GPU tray and motherboard tray updates if the system appears healthy.
B. Update the GPU tray first, then the motherboard tray, and reboot the BMC after all updates are complete.
C. Stop all GPU activity, update and reboot the BMC, update motherboard and tray components, perform a cold reset, and verify completion.
D. Perform a cold reset, stop all GPU activity, update and reboot the BMC, update motherboard and tray components, and verify completion.
Answer: C
Question #10 (Topic: Exam A)
One of the nodes in a cluster is not running as fast as the others and the system administrator needs to check the status of the GPUs on that system.
What command should be used?
What command should be used?
A. iblinkinfo
B. lspci | grep NVIDIA
C. nvidia-smi
D. nvidia-gpu-status
Answer: C