Nebius is looking for a Sr. Solutions Engineer, you will act as the primary technical partner for customers deploying and operating GPU clusters and AI infrastructure on Nebius. You will bridge customer requirements with internal engineering capabilities, ensuring successful deployment, stable operations, and ongoing optimization of complex, high-performance environments
Acting as the main technical interface for customers running workloads on Nebius GPU infrastructure
Supporting customers in deploying, configuring, and tuning GPU-based environments for performance and reliability
Investigating and resolving complex issues spanning hardware, networking, operating systems, and cluster-level behavior
Partnering with internal teams (data center operations, networking, platform engineering) to coordinate and drive resolution of customer-impacting issues
Converting customer requirements into practical architectures, configurations, and execution plans
Identifying opportunities to improve system performance, stability, and overall customer experience
Developing and maintaining technical documentation, including solution patterns, troubleshooting guides, and operational best practices
Contributing to continuous improvement by surfacing recurring issues, gaps, and optimization opportunities to internal teams