WorkaemКарьерная платформа
  • Вакансии
  • Компании
  • Зарплаты
  • Офферы
  • Сервисы
  • Блог
  • Работодателям
Workaem

Карьерная платформа для IT-специалистов: вакансии напрямую с карьерных страниц 300+ компаний, из телеграм-каналов, с международных площадок и от работодателей напрямую. Разбор условий, детектор мёртвых вакансий, AI-инструменты для резюме. Базовые функции бесплатны.

Подпишись, присылаем лучшие вакансии недели
Или читай канал в телеграме
Соискателям
Все вакансииЗа границейУдалёнка в долларахКомпании с РУ основателямиЗарплатыОфферыВозможностиСоветыСоздать резюмеТренировка интервью
По технологиям
Вакансии PythonВакансии JavaScriptВакансии ReactВакансии JavaВакансии GoВакансии Docker
По профессиям
РазработкаДизайнQA / ТестированиеАналитикаProduct / Project ManagerМаркетинг
Работодателям
Разместить вакансиюТарифыБаза кандидатовСвязаться с нами
Кабинет
РегистрацияВойтиЛичный кабинетМои откликиСохранённыеУведомления
Компания
О проектеПредложенияКонтактыБлогКонфиденциальностьУсловия использования
© 2026 Workaem. Все права защищены.КонфиденциальностьУсловияОферта
Made by IT, for IT 💛
Data Engineer, Machine Learning
ВердиктОписаниеИнструментыКомпания
  1. Главная
  2. /
  3. Вакансии
  4. /
  5. Data Engineer, Machine Learning

Sesame·San Francisco·23 июня

Data Engineer, Machine Learning

🏢 ОфисMiddleПолная занятость
Зарплата не указана
Вилки нет, про деньги придётся договариваться с нуля.
Нажмите на сигнал, чтобы увидеть, на чём он основан

Наша компания

Sesame believes in a future where computers are lifelike - with the ability to see, hear, and collaborate with us in ways that feel natural and human. With this vision, we're designing a new kind of computer, focused on making voice agents part of our daily lives. Our team brings together founders from Oculus and Ubiquity6, alongside proven leaders from Meta, Google, and Apple, with deep expertise spanning hardware and software. Join us in shaping a future where computers truly come alive.

Чем предстоит заниматься

We're looking for a Data Engineer to build and maintain the data pipelines that feed Sesame's AI models. You'll collaborate directly with machine learning engineers and researchers — your job is to make sure they have the right data, in the right shape, at the right time to train, evaluate, and ship models
Sesame's data is rich and complex: conversations, voice, sensor signals, and product telemetry. You'll design the systems that take raw, unstructured, multimodal data and turn it into clean, versioned, well-documented datasets that ML teams can trust and build on confidently
This is a deeply technical, infrastructure-focused role — closer to ML engineering than traditional data analytics. You'll be deeply embedded with ML teams, understanding their workflows and building infrastructure that accelerates the full model development lifecycle — from data collection and labeling through training and evaluation
Design and build production data pipelines that prepare conversational, voice, and multimodal data for model training and evaluation
Partner directly with ML engineers to understand data requirements for new models and experiments, and deliver datasets that meet those needs
Build and maintain infrastructure for dataset versioning, lineage tracking, and reproducibility — so any training run can be traced back to its exact data
Develop data quality frameworks that catch issues before they become model quality issues: schema validation, drift detection, and coverage monitoring
Optimise large-scale data processing for cost and performance across Sesame's cloud infrastructure
Build tooling that makes it easy for ML engineers and researchers to discover, explore, and request data independently
Define and enforce data governance and privacy standards, particularly around sensitive conversational and voice data
Contribute to architecture decisions around Sesame's broader data platform as the team and data volume grow

Наши требования

5+ years in data engineering, with meaningful experience supporting ML or AI teams specifically
Strong SQL and Python skills — you'll use both daily
Experience building and operating ETL/ELT pipelines at scale using modern data platforms and tooling
Experience with workflow orchestration systems such as Airflow, Dagster, or Prefect
Hands-on experience with ML data workflows: training data pipelines, dataset versioning, data labeling pipelines, or model evaluation data
A solid understanding of how ML teams work — you don't need to train models; what matters is understanding what makes a good training dataset and why data quality directly affects model performance
Comfort working with unstructured and semi-structured data — audio, text, JSON logs — not just clean relational tables
Strong communication skills. You'll be embedded with ML engineers and need to bridge data systems and model requirements effectively
Vector databases, embedding storage, or feature stores
Data from hardware or embedded systems: telemetry, sensors, real-time streams
Distributed compute frameworks for large-scale data processing such as Ray or Spark
Kubernetes and managed Kubernetes environments such as GKE or EKS
Data privacy frameworks, especially around voice or conversational data
Building internal tooling or self-serve data platforms

Мы предлагаем

401 (k) max employer match: 3.5% of compensation
100% employer-paid health, vision, and dental benefits for you and your dependents
Unlimited PTO and sick time
Flexible spending account with employer matching up to $1,650/year (medical FSA)
Guardian Employee Assistance Program (EAP)
Opportunity to share in the company's success with competitive stock options
S
Sesame
San Francisco

ГрейдMiddle
ЗанятостьПолная занятость
РегионНе Россия
ФорматОфис
ИсточникСкрыто
Опубликовано23 июня
Все вакансии компании

AI-помощник

под эту вакансию
Войди, чтобы AI оценил твоё соответствие вакансии и написал сопроводительное письмо
Мы против мошенников на площадке: если тебя просят заплатить, продиктовать код или установить непонятное приложение, прекращай общение и сразу пиши нам (чат с основателем или форма обратной связи).