Remote Staff Software Engineer - Databases SRE UK Remote

Grafana Labs, Abbots Worthy, Hampshire

Remote Staff Software Engineer - Databases SRE UK Remote

Salary not available. View on company website.

Grafana Labs, Abbots Worthy, Hampshire

  • Full time
  • Permanent
  • Remote working

Posted today, 5 Oct | Get your application in now to be one of the first to apply.

Closing date: Closing date not specified

Job ref: 757e3aa91179469b91e16a1470fb0448

Location ref: Abbots Worthy, Hampshire

Full Job Description

Remote Staff Software Engineer - Databases SRE UK Remote


2026-10-03T00:00:00


Abbots Worthy


Hampshire


GB


SO21 1


Work From Home


2027-01-01T19:02:08.077


Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. If this role excites you, we d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we are looking for candidates from the UK, Sweden, Spain or Germany.


We are looking for a Staff Software Engineer - SRE to help us support our highest value Grafana Cloud customers by increasing the reliability of our Cloud databases that are based on Mimir, Loki, Tempo, and Pyroscope. We provide these databases as a SaaS product from AWS, GCP, and Azure across all regions. The SRE team is embedded within the Mimir, Loki, and Tempo squads and focuses on ensuring that Grafana Cloud s database products deliver exceptional reliability for our highest-SLA customers.


Partner closely with product engineering squads (embedded model) and improve alert quality and reduce noisy escalations. We seek a staff software engineer operating at the intersection of customer needs, production systems, and product engineering.



  • 8+ years engineering experience, 4+ in SRE/CRE/production engineering.

  • Strong preference for those with formal customer reliability engineering experience.

  • Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.).

  • Strong experience with technical leadership, leading a team through projects, mentoring other engineers on the team and serving as a force-multiplier.

  • Experience with one or more programming languages (e.g., Go, Python, Java, etc).

  • Experience with Linux operating systems internals, and some knowledge of networking, cloud storage, and scaling.

  • Experience with calmly and actively participating in blame-free Incident Response, following up on actions, and writing high quality PIRs (Post Incident Reviews, a.k.a. post-mortem documents).

  • Ability to reason about performance, scaling, and failure modes.

  • Comfortable working within an engineering team where individuals are encouraged to have a strong sense of autonomy and self-direction.

  • Ability to partner deeply with product engineering teams.

  • Reviewing and creating SLOs, proactively investigating ways to further reduce budget burn for those SLOs, which can include improvements to monitoring, automation, increasing self-healing, auto-scaling, etc.

  • Collaborating with engineering leaders to help define and influence product strategy, roadmaps, and technical designs.

  • Teaching others about Site Reliability Engineering and communicating best practices to be applied early in development.


In the UK, the base compensation range for this role is £103,958 - £124,750. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable), and other listed benefits.


#LI-Remote


#LI-Remote


Compensation ranges are country-specific.


100% Remote, Global Culture - As a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose.


Innovation-Driven - Autonomy and support to ship great work and try new things.


Open Source Roots - Built on community-driven values that shape how we work.


Career Growth Pathways - Defined opportunities to grow and develop your career.


Passionate People - Join a team of smart, supportive folks who care deeply about what they do.


Balance is Key - We operate a global annual leave policy of 30 days per annum. Three days of your annual leave are reserved for Grafana Shutdown Days to allow the team to disconnect.


We will comply with local legislation where applicable.


Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment, identity or expression, sexual orientation, age, religion or belief, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic protected by law.


We believe that equality and diversity build a strong organization, and we work hard to ensure that is the foundation of our organization as we grow.


Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings.


#s1-Gen

Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. If this role excites
you, we d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we are looking for candidates from the UK, Sweden, Spain or Germany. We are looking for a Staff Software Engineer - SRE to help us support our highest value Grafana Cloud customers by increasing the reliability of our Cloud databases that are based on Mimir, Loki, Tempo, and Pyroscope. We provide these databases as a SaaS product from AWS, GCP, and Azure across all regions. The SRE team is embedded within the Mimir, Loki, and Tempo squads and focuses on ensuring that Grafana Cloud s database products deliver exceptional reliability for our highest-SLA customers. Partner closely with product engineering squads (embedded model) Improve alert quality and reduce noisy escalations We seek a staff software engineer operating at the intersection of customer needs, production systems, and product engineering. 8+ years engineering experience, 4+ in SRE/CRE/production
engineering. Strong preference for those with formal customer reliability engineering experience. Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.). Strong experience with technical leadership, leading a team through projects, mentoring other engineers on the team and serving as a force-multiplier Experience with one or more programming languages (e.g. Go, Python, Java, etc) Experience with Linux operating systems internals, and some knowledge of networking, cloud storage, and scaling. Experience with calmly and actively participating in blame-free Incident Response, following up on actions, and writing high quality PIRs (Post Incident Reviews, a.k.a. post-mortem documents) Ability to reason about performance, scaling, and failure modes Comfortable working within an engineering team where individuals are encouraged to have a strong sense of autonomy and self-direction. Ability to partner
deeply with product engineering teams Reviewing and creating SLOs, proactively investigating ways in which we can further reduce budget burn for those SLOs, which can be self-directed or as the result of learnings from incidents, and may include improvements to monitoring, automation, increasing self-healing, auto-scaling, etc. Collaborating with our Engineering Leaders to help define and influence product strategy, roadmaps and technical designs Teach others about Site Reliability Engineering and communicate best practices to be applied early in development of new features and functionality In the UK, the Base compensation range for this role is £103,958 - £124,750. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here . #LI-Remote #LI-Remote Compensation ranges are country specific. 100% Remote, Global Culture - As a remote-only
company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Innovation-Driven Autonomy and support to ship great work and try new things. Open Source Roots Built on community-driven values that shape how we work. Career Growth Pathways Defined opportunities to grow and develop your career. Passionate People Join a team of smart, supportive folks who care deeply about what they do. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability,
veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. LI-Remote For information about how your personal data is used once you ve applied to a job, check out our .

Direct job link

https://www.jobs24.co.uk/job/remote-staff-software-engineer-databases-sre-uk-127525320