Base Salary
$175k - $210k/yr
Responsibilities
- Design and lead development of backend services, distributed systems, and enterprise features at scale.
- Architect and implement Go backend services with emphasis on correctness, observability, and performance.
- Design APIs and service contracts used by enterprise operators and cloud service providers.
- Drive projects from ideation through development and production operations.
- Improve platform scalability, reliability, security, and multi-tenancy.
- Engage directly with enterprise customers and cloud service providers to translate requirements into engineering solutions.
- Collaborate with Product, UX, and frontend engineers on complete end-to-end solutions.
- Hire, mentor, and develop engineers and improve engineering and operational practices.
- Contribute to open source communities and advocate for customers throughout the development lifecycle.
- Participate in weekday 12-hour-by-5-day and weekend 24-hour-by-2-day on-call rotations.
Requirements
- Deep professional experience writing and operating production services at scale.
- Strong understanding of distributed systems, including replication, consistency models, partitioning, fault tolerance, and scale-related trade-offs.
- Experience designing or operating large-scale, high-traffic, high-availability, or multi-tenant systems.
- Professional experience building and consuming gRPC and protobuf APIs and designing service contracts.
- Strong database skills with PostgreSQL and/or MySQL, including schema design, query optimization, and migrations at scale.
- Experience with large-scale CI/CD systems, build tooling, or continuous delivery pipelines.
- Experience with Kubernetes and containerized deployment environments, including stateful workloads and multi-tenant clusters.
- Experience with observability tooling such as OpenTelemetry, Prometheus metrics, structured logging, and distributed tracing.
- Familiarity with dependency injection patterns such as Google Wire and clean, testable service architecture.
- Preferred experience with TypeScript and React, Grafana’s LGTM+ stack, major cloud or IaaS providers, or large-scale build infrastructure.
Benefits
- 100% remote position for candidates in the US and Canada
- Company-funded usage budget for AI coding assistants
- In-person onboarding
- 30 days of annual leave per year, including 3 Grafana Shutdown Days
- Defined career growth pathways and a transparent, collaborative remote culture
Tech Stack
About Grafana
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions, while getting the visibility they need to run AI systems reliably and at scale. Today, more than 35 million users and 7,000+ customers – including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce – trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,400+ team members across 40+ countries, and we’re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.P. Morgan, CapitalG, and Lead Edge Capital.