WorkaemКарьерная платформа
  • Вакансии
  • Компании
  • Зарплаты
  • Офферы
  • Сервисы
  • Блог
  • Работодателям
Workaem

Карьерная платформа для IT-специалистов: вакансии напрямую с карьерных страниц 300+ компаний, из телеграм-каналов, с международных площадок и от работодателей напрямую. Разбор условий, детектор мёртвых вакансий, AI-инструменты для резюме. Базовые функции бесплатны.

Подпишись, присылаем лучшие вакансии недели
Или читай канал в телеграме
Соискателям
Все вакансииЗа границейУдалёнка в долларахКомпании с РУ основателямиЗарплатыОфферыВозможностиСоветыСоздать резюмеТренировка интервью
По технологиям
Вакансии PythonВакансии JavaScriptВакансии ReactВакансии JavaВакансии GoВакансии Docker
По профессиям
РазработкаДизайнQA / ТестированиеАналитикаProduct / Project ManagerМаркетинг
Работодателям
Разместить вакансиюТарифыБаза кандидатовСвязаться с нами
Кабинет
РегистрацияВойтиЛичный кабинетМои откликиСохранённыеУведомления
Компания
О проектеПредложенияКонтактыБлогКонфиденциальностьУсловия использования
© 2026 Workaem. Все права защищены.КонфиденциальностьУсловияОферта
Made by IT, for IT 💛
Senior Site Reliability Engineer
ВердиктОписаниеИнструментыКомпанияПохожие
  1. Главная
  2. /
  3. Вакансии
  4. /
  5. Senior Site Reliability Engineer

Lodgify·Spain·27 авг.

Senior Site Reliability Engineer

🌍 УдалённоSeniorПолная занятость🌐 Глобал
Зарплата не указана
52
Есть о чём спросить
Навык востребован (python). Но вилки нет, про деньги придётся договариваться с нуля.
Нажмите на сигнал, чтобы увидеть, на чём он основан

Наша компания

Lodgify is a fast-growing scale-up company leading the vacation rental industry. Backed by $30M in funding, our platform empowers property owners and managers worldwide to efficiently manage and grow their business through technology.

О роли

Join Lodgify as a Senior Site Reliability Engineer and help our engineering teams build and operate reliable, observable, scalable, and resilient services by design. Work in the Platform team to strengthen observability, reduce operational toil, improve incident response, define practical SRE standards, and improve the reliability of critical infrastructure and delivery workflows.

Чем предстоит заниматься

Define meaningful SLIs, SLOs, and reliability targets for the platform
Collaborate with the software engineering teams to define and achieve the best practices for software observability, SLIs, SLOs and reliability
Strengthen production readiness by improving service ownership, observability, alerting, runbooks, scaling assumptions, rollback paths, and failure-mode preparedness
Improve the reliability, scalability, and performance of cloud, Kubernetes, and shared infrastructure
Build actionable observability using metrics, logs, traces, and golden signals, with tools such as Datadog, Prometheus, and Grafana
Implement operational and security best practices through guidelines, policies and automation
Reduce alert noise and improve signal quality so teams can detect, understand, and resolve issues quickly
Automate repetitive operational work using Python or other languages
Implement self-service Internal Developer Platform features via APIs and Kubernetes operators
Improve deployment safety, rollbackability, and release observability
Improve reliability of critical stateful systems such as databases, caches, queues, and streaming platforms
Participate in on-call, troubleshoot, and coordinate incident response, and facilitate blameless post-incident reviews that turn into concrete improvements
Execute disaster recovery drills and analyse cloud/platform usage to identify cost and resource-efficiency gains without compromising reliability

Наши требования

You have 7+ years of production experience operating Kubernetes-based platforms and cloud infrastructure
You understand and apply SRE practices: SLIs, SLOs, error budgets, production readiness, incident response, post-incident learning, toil reduction, scalability, capacity planning, high availability, backups, and disaster recovery
You can design and improve observability and alerting for critical systems using metrics, logs, traces, and golden signals, and are comfortable troubleshooting complex distributed systems to identify systemic reliability improvements
You can write maintainable software to automate operational tasks and reduce manual intervention
You have experience with stateful production systems such as relational databases, caches, queues, or streaming platforms
You know how to balance reliability, performance, cost, and delivery speed pragmatically
You are comfortable working in a transitional environment where SRE practices are being introduced while critical infrastructure and delivery systems still need hands-on reliability support
You collaborate effectively with Engineering, Platform, Security, and Product stakeholders
You communicate clearly, document well, and enjoy coaching teams toward stronger production ownership
You model initiative and accountability, raising risks early and driving improvements through to completion

Стек

KubernetesCloud infrastructureDatadogPrometheusGrafanaPython

Мы предлагаем

Remote Flexibility: The freedom to work from home any day that works for you
Time to Recharge: 25 working days of paid vacation and Jornada Intensiva in August
Alan Health Insurance: Premium health, dental, and mental health support via Alan. Pre-existing conditions are covered
Meal Perk: €150/month allowance on your Alan card + 50% off Ametller Origen prepared dishes at the office
Tax-Free Savings: Increase your take-home pay by using Flexible Remuneration for extra meal costs (up to €70/mo) and public transport (up to €136/mo)
Home Office Gear: We provide a table, ergonomic chair, and monitor for your home setup
Language Learning: Free Spanish classes
Referrals: Cash rewards for bringing in new talent
Social Life: Daily office breakfast and monthly team events
Dynamic Hub: A high-energy, inclusive environment designed for collaboration and connection with a team that represents over 60 countries

Дополнительно

Benefits offered may differ based on the type of contract that is issued
All applications and CVs must be submitted in English

Технологии и навыки

engineer
python
cloud
security
kubernetes
L
Lodgify
Spain

ГрейдSenior
ЗанятостьПолная занятость
РегионИспания
ФорматУдалённо
ИсточникСкрыто
Опубликовано27 авг.
Все вакансии компании

AI-помощник

под эту вакансию
Войди, чтобы AI оценил твоё соответствие вакансии и написал сопроводительное письмо

Похожие вакансии

Machine Learning Engineer III, Routing CostMapboxAlliances Field EngineerCanonicalCloud Field EngineerProДоступна только зарегистрированнымDedicated Linux Desktop & Devices Support EngineerProДоступна только зарегистрированным
Мы против мошенников на площадке: если тебя просят заплатить, продиктовать код или установить непонятное приложение, прекращай общение и сразу пиши нам (чат с основателем или форма обратной связи).