Datacenter Hardware Technician
The job description
Tech stack. Server component replacement (drives, DIMMs, PSUs, system boards), firmware update campaigns, hardware diagnostics suites, ESD and FRU procedures, asset management systems, quality inspection, torque specifications
About the role
You will specialize in server hardware repair at a large-scale datacenter operator running a heterogeneous fleet from multiple vendors and generations. Your craft is the physical server itself: precise component replacement, disciplined firmware management, and quality verification that returns each machine to the fleet in better documented condition than you found it. Your first-time-fix rate, your rework rate, and your throughput are the metrics that define your contribution, and all three must move in the right direction together. You will develop expertise across the full fleet portfolio, becoming the go-to specialist for the trickiest repair scenarios and unusual failure modes. Your repair data will help reliability engineering distinguish between random failures and systematic design problems.
What you will achieve
- Complete hardware repairs with first-time-fix rates above 92 percent and rework rates below 2 percent, verified by post-repair diagnostic suites run on every serviced machine.
- Execute firmware update campaigns across hundreds of servers per week with zero bricking incidents, through disciplined staging, verification, and rollback procedures.
- Reduce parts waste and cost by diagnosing accurately to the FRU level, avoiding unnecessary full-board or full-server replacements that inflate spares consumption.
- Maintain repair throughput meeting daily targets without sacrificing quality gates, safety procedures, or documentation completeness.
- Document repair findings with failure detail that feeds reliability engineering, giving design and procurement teams real fleet failure data instead of anecdotes.
What you will bring
Must-haves
- 2 to 5 years of hands-on server or computer hardware repair experience at meaningful volume.
- Detailed knowledge of server components: drives, memory, CPUs, power supplies, RAID controllers, NICs, and their failure signatures.
- Proficiency with vendor hardware diagnostic tools and firmware update utilities across at least one major platform.
- Strict ESD discipline and component handling practice, including proper packaging for returns.
- Ability to work efficiently to throughput targets while maintaining quality without cutting corners.
- High school diploma or equivalent; technical certifications valued.
Nice-to-haves
- OEM certifications (Dell, HPE, Supermicro, Lenovo) for server hardware service.
- Experience with GPU server hardware, its thermal requirements, and its unique failure modes.
- Familiarity with liquid-cooled server service procedures and coolant handling safety.
- Knowledge of secure data handling for drives removed during repair.
Google
Meta
Microsoft
Amazon
Equinix
Digital Realty