5+ years of professional experience in system-level software development (focused on performance optimization, low-level programming)
3+ years of hands-on experience with Linux systems (administration, troubleshooting, and performance tuning)
In-depth understanding of server architecture, including PCIe devices, NICs, Linux OS/Kernel, and high-performance computing (HPC) systems
Strong proficiency in one or more performance-oriented programming languages (C/C++, Go, Python)
Experience with GPU end-to-end testing in a cluster environment using InfiniBand networking
Proven track record of analyzing and optimizing the performance of HPC workloads (e.g., simulations, data analysis, AI/ML workloads)
Familiarity with RDMA, RoCE, and InfiniBand protocols for high-performance communication
Background in Software-Defined Networking (SDN) and experience with HPC cluster networking
Understanding of QEMU/KVM virtualization and managing virtualized environments
Experience with deep learning frameworks such as PyTorch and TensorFlow, and their integration with HPC systems
Familiarity with collective communication libraries like MPI and NCCL for distributed computing
We offer competitive salaries ranging from $170k-$300k + equity based on your experience
We conduct coding interviews as part of the process