Grafana Labs

Grafana Labs

Backend Engineer – Platform – Stacks | Ireland | Remote

UK, Spain, Germany, Sweden, IrelandRemotePosted Today
Full TimeSeniorRemoteIE

See how this job matches your profile

Sign in for an AI-powered fit score, breakdown, and a tailored resume.

Sign in

Job Description

Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telem

Key Highlights

  • Design, build, and operate reconciliation systems, including the SSS backend, to track desired stack state, detect and repair drift across stack templates, grafana.com state, Hosted Grafana, and actual customer stack configuration
  • Collaborate across SSS, grafana.com, and deployment configurations to ensure stack lifecycle workflows remain reliable, observable, and resilient
  • Improve operational efficiency by reducing deployment complexity (e.g., aiming for single PR regional SSS deployment) and contributing to the Stack Config Reconciliation project
  • Manage rollout mechanisms for provisioned plugins, dashboards, data sources, Grafana versions, release channels, and stack-level configuration
  • Support new region and cluster rollouts, including the operational paths required to bring stacks online safely in new Grafana Cloud regions

Qualifications

Required Qualifications

  • You have at least 1 year of fully remote work experience
  • You have some experience working on a SaaS platform and are familiar with common distributed systems concepts (e.g., scalability, multi-tenancy, HA).
  • Have professional experience with Golang and be willing to work across both backend service and application code
  • Care deeply about developer and user experience and the quality of the products that you work on
  • Have some experience contributing to the delivery of projects, from initial brainstorming to shipping a product to the customer.
  • You write clean, well-tested software that other engineers can understand, operate, and maintain
  • Can take on well-defined tasks, break them down, and execute iteratively to deliver working solutions and gather feedback.
  • You are willing to collaborate across teams and ensure your work is aligned with the needs of other squads and external stakeholders.
  • Familiarity with Kubernetes in AWS, GCP, or Azure, and exposure to infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.).
  • Experience participating in blameless incident response and contributing to post-incident reviews.

Preferred Qualifications

  • Experience with TypeScript/Node.js
  • Experience with Kubernetes control-plane patterns, operators, reconcilers, or desired-state systems
  • Experience with Jsonnet/Tanka, Terraform, Flux, Argo, or similar deployment/configuration tooling
  • Experience working on SaaS provisioning, tenancy, regional expansion, plugin rollout, or customer lifecycle systems
  • Experience with incident response involving configuration drift, partial failure, or cross-service state mismatch

Skills & Technologies

AWSAzureGCPGolangKubernetesTerraformTypeScriptNode.js

Interested in this role?

Sign in or create a free account to see how this job matches your skills, apply with one click, and let our AI tailor your resume.

Sign in to apply
AI-powered resume optimization
Save and track your applications

Job Details

Employment Type

Full Time

Experience Level

Senior

Location

UK, Spain, Germany, Sweden, Ireland

Work Mode

Remote

Posted

Today

Country

IE