WorkaemКарьерная платформа
  • Вакансии
  • Компании
  • Зарплаты
  • Офферы
  • Сервисы
  • Блог
  • Работодателям
Workaem

Карьерная платформа для IT-специалистов: вакансии напрямую с карьерных страниц 300+ компаний, из телеграм-каналов, с международных площадок и от работодателей напрямую. Разбор условий, детектор мёртвых вакансий, AI-инструменты для резюме. Базовые функции бесплатны.

Подпишись, присылаем лучшие вакансии недели
Или читай канал в телеграме
Соискателям
Все вакансииЗа границейУдалёнка в долларахКомпании с РУ основателямиЗарплатыОфферыВозможностиСоветыСоздать резюмеТренировка интервью
По технологиям
Вакансии PythonВакансии JavaScriptВакансии ReactВакансии JavaВакансии GoВакансии Docker
По профессиям
РазработкаДизайнQA / ТестированиеАналитикаProduct / Project ManagerМаркетинг
Работодателям
Разместить вакансиюТарифыБаза кандидатовСвязаться с нами
Кабинет
РегистрацияВойтиЛичный кабинетМои откликиСохранённыеУведомления
Компания
О проектеПредложенияКонтактыБлогКонфиденциальностьУсловия использования
© 2026 Workaem. Все права защищены.КонфиденциальностьУсловияОферта
Made by IT, for IT 💛
Principal Network Engineer
ВердиктОписаниеИнструментыКомпания
  1. Главная
  2. /
  3. Вакансии
  4. /
  5. Principal Network Engineer

TensorWave·31 июля

Principal Network Engineer

🌍 УдалённоLeadПолная занятость🌐 Глобал
Зарплата не указана
Вилки нет, про деньги придётся договариваться с нуля.
Нажмите на сигнал, чтобы увидеть, на чём он основан

Наша компания

Our mission is simple: deliver seamless, secure, reliable, and resilient AI compute at scale. We've built a versatile cloud platform that eliminates infrastructure barriers, empowering builders to focus on innovation instead of fighting their stack. Because breakthrough AI should move at the speed of ideas, not infrastructure.

Чем предстоит заниматься

We’re seeking a Principal Network Engineer (L7) focused on owning and evolving large-scale, RoCEv2 data center networks powering next generation AI and ML infrastructure
You’ll work closely with our network architect and infrastructure leadership to define how the network is designed, implemented, and operated at scale, keeping over 8,000 GPUs burring today and scaling to cluster sizes reaching over 100,000 GPUs. You will be responsible for the architectural decisions that determine performance, reliability, and operational sanity at scale
You’ll remain hands-on with high-speed optics, switching, routing, and congestion management in production clusters, while also setting the standards, patterns, and tooling other engineers build and operate against
As a Principal Engineer, you own end-to-end network architecture, make high-impact design decisions, and set technical direction across teams, with clear examples of systems you’ve defined and scaled
Define, evolve, and standardize large-scale RoCEv2 data center networks supporting AI and ML clusters from thousands to 100,000+ GPUs
Set and validate congestion management strategy across RDMA fabrics, including PFC, ECN, and DCQCN, based on real production behavior
Establish automation, validation, and observability patterns that prevent misconfiguration and eliminate manual operational work
Act as the technical escalation point for complex failures, scaling limits, and architectural tradeoffs in always-on, multi-tenant environments

Наши требования

Bachelor’s degree in Computer Science, Electrical Engineering, or a related technical field, or equivalent practical experience
Deep experience designing and operating RDMA and RoCEv2 networks in large-scale production environments supporting AI or HPC workloads
Expert-level knowledge of switching hardware and their NOS, such as Arista, Juniper, and custom solutions using SONiC, including high-speed Ethernet fabrics
Proven hands-on experience with congestion management and performance tuning using PFC, ECN, and DCQCN
Strong experience with high-speed optics and cabling including 400G, 800G, and AEC, AOC, DAC, and structured cabling at scale
Strong automation mindset, with experience using Python, Ansible, Terraform, Git, and production observability tooling
We’re looking for engineers who operate comfortably at scale, make hard calls with incomplete data, and take responsibility for systems that must work under sustained load. The solutions that work on a handful of devices will not work at Exascale

Мы предлагаем

Stock Options
100% paid Medical, Dental, and Vision insurance for Employees
Company Health Savings Account Contributions
100% paid Short Term and Long Term Disability Insurance for Employees
Life and Voluntary Supplemental Insurance Options
Other Insurance Options, such as Pet & Legal Insurance
Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support
Flexible Spending Account
401(k)
Employee Assistance Program
Flexible PTO
Paid Holidays
Parental Leave
Other In-Office Perks
Employment Eligibility
All offers of employment are contingent upon verification of identity and authorization to work in the United States, as required by law
Background Checks
Where permitted by law, employment may be contingent upon the successful completion of a job-related background check
Data Privacy Notice
By submitting an application, you acknowledge that TensorWave may collect, use, and retain your personal information for recruiting and employment-related purposes in accordance with applicable data privacy laws
T
TensorWave

ГрейдLead
ЗанятостьПолная занятость
РегионНе Россия
ФорматУдалённо
ИсточникСкрыто
Опубликовано31 июля
Все вакансии компании

AI-помощник

под эту вакансию
Войди, чтобы AI оценил твоё соответствие вакансии и написал сопроводительное письмо
Мы против мошенников на площадке: если тебя просят заплатить, продиктовать код или установить непонятное приложение, прекращай общение и сразу пиши нам (чат с основателем или форма обратной связи).