Grafana Cloud is our composable observability platform that integrates metrics, logs, and traces with Grafana. It allows our customers to leverage the best open source observability software – including Prometheus, Mimir, Loki, and Tempo – without the overhead of installing, maintaining and scaling their own observability stack
The Databases team owns the telemetry databases that are Mimir for metrics, Loki for logs, Tempo for traces, and Pyroscope for profiles. Our databases are OSS projects that we also offer as a Cloud service supporting Grafana Cloud, and as an on-premise solution. They are multi-tenant distributed systems implemented in Go and running on Kubernetes across all major Cloud service providers (GCP, Azure, AWS). We also have engineers working on Prometheus, Grafana Agent, Mimir proxies, and OpenTelemetry
As a company we are remote-first and global, we embrace people of different experiences and backgrounds to build diverse teams where every person brings a new perspective to the software. Our tech stack is mostly made up of services written in Go, running on multiple Kubernetes clusters that leverage Cloud object storage
We invest heavily in developer productivity. You can use modern AI coding assistants as part of your daily workflow (your choice of tools, within security guidelines), backed by a company-funded usage budget so you can iterate quickly without unnecessary friction. We encourage pragmatic AI-assisted development: faster prototyping, test generation, refactors, documentation, and incident follow-ups—always paired with strong code review and quality standards. You’ll also have access to frontier models (e.g., GPT-Codex 5/3, Claude Opus 4.6, Gemini 3 Pro)
Take an active role in influencing our roadmap and your own career objectives
Work with your team to deliver new features, then use the results to iterate and improve
Drive projects from initial ideation all the way to operations once it is in the hands of customers
Embrace our open-source culture and contribute to other projects that may not directly fall within your team’s scope
Design, build, operate, and maintain critical systems, owning the reliability, performance, and availability
Be a part of your team’s on-call rotations and take ownership of the services you’re running
Supporting other team members, participate in design discussions and collaborate with the team
Learn new skills by gaining a deeper understanding of our cloud product and our customers and getting to know the codebase of a large distributed system
As we are remote-first and our engineering organization is largely remote, we provide guidance and meet regularly using video calls, so an independent attitude and good communication skills are a must