Data Center Technician

  • $60,000–$100,000 Per Year
  • Full-time

Highlights

High-Value Skills: High-Performance Computing (HPC) clusters (Slurm, Bright Cluster Manager), liquid cooling, dense rack layout, networking protocols (TCP/IP, DNS, NFS, SSL). Core Specialty: GPU & Server Hardware Break/Fix, Linux/Windows OS Administration, DCIM Tooling, Scripting (Python/Shell/Ansible).

Numbers & Facts

LocationNV
Job TypeFull-time
Salary$60,000–$100,000 Per Year

Description

Job Summary: 

Position Title: Data Center Technician / Engineer
Location: Reno, NV (5 Days Onsite)
Duration: Long Term/ Fulltime
Target Experience: 5 to 8 Years
Core Specialty: GPU & Server Hardware Break/Fix, Linux/Windows OS Administration, DCIM Tooling, Scripting (Python/Shell/Ansible)
High-Value Skills: High-Performance Computing (HPC) clusters (Slurm, Bright Cluster Manager), liquid cooling, dense rack layout, networking protocols (TCP/IP, DNS, NFS, SSL)

Job Responsibilities

Key Responsibilities:

Hardware & Compute Farm Management: Maintain a high-performing compute farm of builders, packagers, testers, and core server infrastructure.
Server & GPU Break/Fix: Perform hands-on troubleshooting and replacement for PCBs, GPUs, power supplies, memory, and high-density compute nodes.
Automation & Scripting: Use Shell, Python, or Ansible to automate recurring tasks, run operational scripts, and manage DCIM tooling (e.g., Nautobot).
Cross-Functional Operations: Collaborate with system architects, software developers, and QA engineers to debug hardware/software edge cases and meet availability SLAs.
Process Documentation: Author Standard Operating Procedures (SOPs), collect key operational metrics, and manage system recovery efforts.
 
 

Qualifications

Required Qualifications & Skills:

Associate’s or Bachelor’s degree in a technical major (or equivalent hands-on experience).
5 to 8 years of direct experience in data center environments or large engineering labs.
Strong operating system administration across Linux, Windows, and macOS.
Hands-on scripting proficiency with Python, Shell, or Ansible.
Working knowledge of network protocols: TCP/IP, DNS, NFS, SSL.
Direct experience with DCIM tools (Nautobot or similar inventory/rack management systems).
 

Preferred / Standout Skills:

Experience managing HPC clusters using Slurm or Bright Cluster Manager (BCM).
Knowledge of dense server infrastructure, including liquid cooling systems.
Network certifications such as CCNA or equivalent.

Skills

  • HVAC
  • Network Operations Center
  • Hardware Installation
  • HPC
  • GPU (Graphics Processing Unit)
  • Linux Operating System
  • Linux Administration
  • Network Administration/Management
  • Information Technology & Information Systems
  • Network System Hardware
  • Maintenance Services
  • Network Switching

Similar Jobs

See more jobs