NCP-AII Question 19
Select 3You are tasked with replacing a faulty GPU in a server within your AI infrastructure. To ensure proper reinstallation and minimize downtime, which steps should you take during this process?
- A
Power off the server and disconnect it from the power source before beginning the replacement.
- B
Use an anti-static wrist strap to prevent damage to components during handling.
- C
Remove the faulty GPU and immediately install its replacement without verifying the compatibility of the new GPU with the server.
- D
Inspect the system logs and firmware settings after installation to ensure the new GPU is recognized properly.
- E
Skip running stress tests after replacing the GPU, as they are not necessary for hardware replacements.
Show answer and explanation
Correct answers: A, B, D
Explanation
Replacing a faulty GPU requires strict adherence to safety protocols, such as powering off the server and using anti-static measures. Additionally, verifying compatibility, checking system logs, and running stress tests are critical to ensure the successful integration and operation of the new GPU.
- A. Correct.
Powering off the server and disconnecting it from the power source is essential to ensure safety and prevent electrical damage during hardware replacement.
- B. Correct.
Using an anti-static wrist strap helps to prevent electrostatic discharge, which can damage sensitive components like GPUs during handling.
- C. Incorrect.
Skipping the verification of compatibility can lead to system failures or improper functionality, so this step is incorrect.
- D. Correct.
Inspecting system logs and ensuring that the new GPU is detected by the firmware is a critical step to confirm the replacement was successful.
- E. Incorrect.
Stress tests are crucial to ensure the new GPU is functioning correctly under load, so skipping them is not recommended.