5+ years of experience in data center operations, infrastructure support, systems administration, hardware support, cloud infrastructure, or similar technical environments
Hands-on experience troubleshooting bare metal servers, Linux systems, hardware components, and networking issues in production environments
Intermediate Linux command-line proficiency, including service validation, log review, process inspection, network configuration checks, and system diagnostics
Experience with server hardware diagnostics, component replacement, firmware updates, BIOS configuration, driver troubleshooting, or BMC tools such as iDRAC, iLO, IPMI, Redfish, or similar
Strong networking fundamentals, including TCP/IP, VLANs, DNS, DHCP, routing/switching concepts, optics, cabling, and link-level troubleshooting
Experience participating in incident response, escalation handling, root cause analysis, or operational recovery in high-availability environments
Strong communication and documentation skills, with the ability to work cross-functionally with data center, infrastructure, network, systems, and engineering teams
Experience with NVIDIA GPUs, GPU servers, HPC, AI infrastructure, or high-density compute platforms
Experience with Supermicro, Dell, HPE, Lenovo, Cisco UCS, Arista, Cisco, Juniper, or similar infrastructure platforms
Familiarity with InfiniBand, RoCE, MPO/MTP fiber, high-speed Ethernet, or clustered compute environments
Experience with Python, Bash, Ansible, Terraform, PowerShell, or similar automation tools
Familiarity with Jira, Confluence, ServiceNow, Grafana, Prometheus, Kubernetes, Docker, or similar operational tools
We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law