Datacenter Technician
The job description
Tech stack. Server hardware troubleshooting, ticketing systems, Linux basics, hardware diagnostics suites, ESD procedures, asset tagging and tracking, KVM and remote management (iDRAC/iLO/IPMI), fiber and copper cabling basics, power distribution awareness
About the role
You will keep a hyperscale datacenter running as part of the team responsible for tens of thousands of servers supporting cloud customers who notice every minute of downtime. When hardware fails, and it fails constantly at this scale, you are the response: diagnosing faults accurately, replacing components efficiently, and returning capacity to the fleet fast. Your work directly determines fleet availability percentages, and your discipline with procedures, documentation, and safety keeps a massive operation consistent across hundreds of technicians and three shifts. You will work across all shifts in a team that takes pride in fleet availability numbers most companies never achieve. Your growth path into senior roles is clear and earned through demonstrated diagnostic skill and reliability.
What you will achieve
- Resolve hardware break/fix tickets within SLA targets consistently, diagnosing faults accurately to the failed component and completing repairs with first-time-fix rates above 90 percent.
- Reduce mean time to repair in your area by following structured diagnostics: isolating faults to the FRU level through logs and testing before replacing parts, not through trial-and-error swaps.
- Maintain perfect safety and ESD compliance across every task and every shift, protecting both personnel and sensitive hardware with zero procedural violations.
- Keep asset and ticket documentation completely accurate so fleet health dashboards reflect reality and capacity planning decisions rest on trustworthy data.
- Identify repeat failure patterns in your area (bad batches, common modes, environmental contributors) and escalate them with data, contributing to fleet-wide reliability improvements.
What you will bring
Must-haves
- 2 to 5 years of hands-on hardware experience: servers, networking gear, storage systems, or similar IT infrastructure.
- Ability to diagnose hardware faults systematically using system logs, built-in diagnostics, and physical inspection rather than guessing.
- Comfort working in a datacenter environment: raised floors, hot aisles, noise, physical security protocols, and shift work.
- Basic Linux command-line skills for checking system state, reading logs, and verifying repairs.
- Strong procedural discipline and clear, complete written communication in ticketing systems.
- High school diploma or equivalent; technical certifications valued.
Nice-to-haves
- CompTIA Server+, A+, or Network+ certification.
- Experience with remote management interfaces (iDRAC, iLO, IPMI) for out-of-band diagnostics.
- Familiarity with lean, 5S, or continuous improvement practices in operations environments.
- Basic scripting ability for log parsing or repetitive operational tasks.
Google
Meta
Microsoft
Amazon
Equinix
Digital Realty