We are looking for a Senior Cloud Engineer to join our Cloud Engineering team. This team is tasked with ensuring scalability, reliability, and security for all backend services and keeping the SDLC running smoothly for engineers across the company
As a senior member of the team, you will own meaningful pieces of our cloud platform end-to-end — from design through deployment and operation — while helping shape the standards, tooling, and golden templates that the rest of the engineering organization builds on. You will balance deep, hands-on technical ownership with mentoring teammates and influencing the technical direction of the systems that power a large-scale mobile marketplace
Scale & Impact: You will work on systems that drive the end-user experience and support every technical team across the organization, operating across a multi-account AWS and GCP footprint
Ownership: You will own critical infrastructure and platform capabilities end-to-end, making architectural decisions that affect reliability, security, and cost at scale
Force Multiplier: The templates, pipelines, and tooling you build raise the productivity and safety of every engineer at OfferUp
Emerging Technology: You will help bring new capabilities — including AI/ML developer tooling and infrastructure — into the hands of our engineers
In this role, you will design, build, and operate the cloud infrastructure, platform tooling, and CI/CD systems that underpin OfferUp's core applications. Day to day, that means you will
Build cloud infrastructure as code — design, implement, and maintain AWS and GCP infrastructure with Terraform, optimizing for reliability, security, and cost
Own our Kubernetes & GitOps platform — run and evolve EKS/GKE (cluster upgrades, Envoy/Gloo networking, Helm, ArgoCD delivery)
Keep the SDLC fast and safe — build CI/CD pipelines and developer workflows (GitHub Actions, monorepo tooling, JFrog Artifactory), and maintain the backend “golden templates” and shared libraries the wider org builds on
Lead platform migrations end-to-end — plan, execute, and de-risk observability, data-store, Kubernetes, and multi-account migrations, including rollback strategy
Drive operational excellence & security — own observability in Datadog and on-call, remediate CVEs, manage secrets and access controls (Cloudflare/ZTNA, Okta, SSO, Workload Identity Federation), and optimize cloud cost
Multiply the team — enable AI/ML developer tooling (Amazon Bedrock, Claude, AI-assisted workflows), lead design and code reviews, set technical standards, mentor engineers, and partner with other teams