Remote Staff Software Engineer - Databases SRE UK Remote
Grafana Labs, Geddington, Northamptonshire
Remote Staff Software Engineer - Databases SRE UK Remote
Salary not available. View on company website.
Grafana Labs, Geddington, Northamptonshire
- Full time
- Permanent
- Remote working
Posted today, 5 Oct | Get your application in now to be one of the first to apply.
Closing date: Closing date not specified
Job ref: 0e835819895e408f95aa04b9ac3fffcd
Location ref: Geddington, Northamptonshire
Full Job Description
Remote Staff Software Engineer - Databases SRE UK Remote
2026-10-03T00:00:00
Foxhall
Northamptonshire
GB
NN14 1
Work From Home
2027-01-01T23:02:34.973
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost.
We are a 100% remote company with 1,600+ team members across 40+ countries, and we are backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J. Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. If this role excites you, we’d love you to raise your hand for what could be a truly career-defining opportunity.
This is a remote opportunity and we are looking for candidates from the UK, Sweden, Spain or Germany. We are looking for a Staff Software Engineer - SRE to help us support our highest value Grafana Cloud customers by increasing the reliability of our Cloud databases that are based on Mimir, Loki, Tempo, and Pyroscope. We provide these databases as a SaaS product from AWS, GCP, and Azure across all regions. The SRE team is embedded within the Mimir, Loki, and Tempo squads and focuses on ensuring that Grafana Cloud’s database products deliver exceptional reliability for our highest-SLA customers. Partner closely with product engineering squads (embedded model) to improve alert quality and reduce noisy escalations. We seek a staff software engineer operating at the intersection of customer needs, production systems, and product engineering.
- 8+ years engineering experience, 4+ in SRE/CRE/production engineering.
- Strong preference for those with formal customer reliability engineering experience.
- Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.).
- Strong experience with technical leadership, leading a team through projects, mentoring other engineers on the team and serving as a force-multiplier.
- Experience with one or more programming languages (e.g., Go, Python, Java, etc).
- Experience with Linux operating systems internals, and some knowledge of networking, cloud storage, and scaling.
- Experience with calmly and actively participating in blame-free Incident Response, following up on actions, and writing high-quality PIRs (Post Incident Reviews, a.k.a. post-mortem documents).
- Ability to reason about performance, scaling, and failure modes.
- Comfortable working within an engineering team where individuals are encouraged to have a strong sense of autonomy and self-direction.
- Ability to partner deeply with product engineering teams.
- Reviewing and creating SLOs, proactively investigating ways to further reduce budget burn for those SLOs, which can be self-directed or as a result of learnings from incidents, and may include improvements to monitoring, automation, increasing self-healing, auto-scaling, etc.
- Collaborating with engineering leaders to help define and influence product strategy, roadmaps, and technical designs.
- Teach others about Site Reliability Engineering and communicate best practices early in the development of new features and functionality.
In the UK, the base compensation range for this role is £103,958 - £124,750. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable), and other listed benefits.
#LI-Remote #LI-Remote
Compensation ranges are country-specific. 100% Remote, with a Global Culture — as a remote-only company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Innovation-driven environment that supports shipping great work and trying new things. Open source roots built on community-driven values. Career growth pathways with opportunities to develop. Passionate team of supportive individuals. Balance with a global annual leave policy of 30 days, including 3 days reserved for Grafana Shutdown Days.
Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment, sexual orientation, age, religion, disability, veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or other protected characteristics. We value diversity and inclusivity as core to our organization’s strength.
Grafana Labs may utilize AI tools in its recruitment process to assist in matching applicant data to job postings. For more information about how your data is used, please refer to our privacy policies.
#s1-Gen
Grafana Labs, the company behind the open observability cloud, is founded on the principles of open source, open standards, open ecosystems, and open culture. Grafana Cloud, our fully managed observability platform, is flexible and built for scale. With Grafana Cloud's actually useful AI, organizations can see, understand, and act on all their disparate data to move at the speed of their ambitions. Today, more than 35 million users and 7,000+ customers including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce trust Grafana Labs to ensure reliability of their applications and systems, resolve incidents quickly, and optimize their telemetry to reduce noise and cost. We are a 100% remote company with 1,600+ team members across 40+ countries, and we re backed by leading investors including Lightspeed Venture Partners, Sequoia Capital, GIC, Coatue, J.Our team thrives in an innovation-driven environment where transparency, autonomy, and trust fuel everything we do. If this role excites
you, we d love you to raise your hand for what could be a truly career-defining opportunity. This is a remote opportunity and we are looking for candidates from the UK, Sweden, Spain or Germany. We are looking for a Staff Software Engineer - SRE to help us support our highest value Grafana Cloud customers by increasing the reliability of our Cloud databases that are based on Mimir, Loki, Tempo, and Pyroscope. We provide these databases as a SaaS product from AWS, GCP, and Azure across all regions. The SRE team is embedded within the Mimir, Loki, and Tempo squads and focuses on ensuring that Grafana Cloud s database products deliver exceptional reliability for our highest-SLA customers. Partner closely with product engineering squads (embedded model) Improve alert quality and reduce noisy escalations We seek a staff software engineer operating at the intersection of customer needs, production systems, and product engineering. 8+ years engineering experience, 4+ in SRE/CRE/production
engineering. Strong preference for those with formal customer reliability engineering experience. Strong Kubernetes experience in AWS, GCP, or Azure, and familiarity with infrastructure-as-code tooling (Helm, Terraform, Jsonnet, etc.). Strong experience with technical leadership, leading a team through projects, mentoring other engineers on the team and serving as a force-multiplier Experience with one or more programming languages (e.g. Go, Python, Java, etc) Experience with Linux operating systems internals, and some knowledge of networking, cloud storage, and scaling. Experience with calmly and actively participating in blame-free Incident Response, following up on actions, and writing high quality PIRs (Post Incident Reviews, a.k.a. post-mortem documents) Ability to reason about performance, scaling, and failure modes Comfortable working within an engineering team where individuals are encouraged to have a strong sense of autonomy and self-direction. Ability to partner
deeply with product engineering teams Reviewing and creating SLOs, proactively investigating ways in which we can further reduce budget burn for those SLOs, which can be self-directed or as the result of learnings from incidents, and may include improvements to monitoring, automation, increasing self-healing, auto-scaling, etc. Collaborating with our Engineering Leaders to help define and influence product strategy, roadmaps and technical designs Teach others about Site Reliability Engineering and communicate best practices to be applied early in development of new features and functionality In the UK, the Base compensation range for this role is £103,958 - £124,750. Actual compensation may vary based on level, experience, and skillset as assessed in the interview process. Benefits include equity, bonus (if applicable) and other benefits listed here . #LI-Remote #LI-Remote Compensation ranges are country specific. 100% Remote, Global Culture - As a remote-only
company, we bring together talent from around the world, united by a culture of collaboration and shared purpose. Innovation-Driven Autonomy and support to ship great work and try new things. Open Source Roots Built on community-driven values that shape how we work. Career Growth Pathways Defined opportunities to grow and develop your career. Passionate People Join a team of smart, supportive folks who care deeply about what they do. Balance is Key - We operate a global annual leave policy of 30 days per annum. 3 days of your annual leave entitlement are reserved for Grafana Shutdown Days to allow the team to really disconnect. We will comply with local legislation where applicable. Equal Opportunity Employer: Grafana Labs is an equal opportunities employer. We welcome applications from everyone regardless of race, colour, nationality, origin, caste, sex, gender reassignment identity or expression, sexual orientation, age, religion or belief, disability,
veteran status, genetic information, pregnancy, maternity, marital, family or carer status, or any other characteristic which is protected by local law. We believe that equality and diversity build a strong organisation, and we work hard to ensure that is the foundation of our organisation as we grow. Grafana Labs may utilize AI tools in its recruitment process to assist in matching information provided in CVs to job postings. LI-Remote For information about how your personal data is used once you ve applied to a job, check out our .