NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Site Reliability Engineer in United Kingdom
2 months ago
Apply with autofill
Apply with autofill
Brevanhoward·2 months ago
2 months ago

Senior Site Reliability Engineer

London (82), United KingdomMid · 5-8 yearsSite Reliability Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at brevanhoward

How you compare FREE

?
Your scoreYour score: not yet known
→
57
Top 10%Top 10%: 57 out of 100

Top 10% of NextRaise users matched against Site Reliability Engineer roles in United Kingdom.

PDF or DOCX · no account needed

Apply faster with autofill FREEThe NextRaise extension autofills your application in one click.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

About this role

Senior Site Reliability Engineer (SRE) - GCP/Kubernetes

About the Role 

We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership.

The successful candidate will split their time between providing expert operational support for our critical systems and leading exciting new infrastructure projects. Our mindset is to get the right person not the person with the skills that match our stack. It is important to be able to foresee problems before they show up and create solutions that mitigate them. If you enjoy a challenging environment, implementing "infrastructure as code" principles, and directly seeing the impact of your work, this is the place for you.

What You Will Do

  • Design & Build: Architect, deploy, and maintain highly scalable and reliable infrastructure on Google Cloud Platform (GCP) using Kubernetes and Infrastructure-as-Code tools.
  • Automation: Champion automation across the entire software development lifecycle (SDLC), utilizing IaC, Python and Bash to reduce toil and improve operational efficiency.
  • Infrastructure-as-Code (IaC): Own and evolve our declarative infrastructure using Terraform for cloud resources and Helm for Kubernetes application deployment.
  • Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification.
  • Reliability & Performance: Define, measure, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Participate in on-call rotation (if applicable) and lead post-incident reviews to drive continuous improvement.
  • Collaboration: Work closely with software development teams to provide expert guidance on deployment strategies, scalability concerns, and cloud-native best practices.
  • Ownership: Take full ownership of projects from inception through to production operation, including documentation and knowledge transfer.

Required Experience & Skills

Core Technical Stack

  • Cloud Platform: 4+ years of hands-on experience with Google Cloud Platform (GCP) (or similar cloud infrastructure).
  • Container Orchestration: Expert-level proficiency in managing, scaling, and troubleshooting production Kubernetes environments.
  • Infrastructure-as-Code: Deep expertise in Terraform for managing cloud and Kubernetes resources.
  • Deployment: Strong experience with Helm for packaging and deploying applications on Kubernetes.
  • Scripting/Programming: Proficient in at least one major programming language, preferably Python, for automation and tool development.

Tooling & Concepts

  • CI/CD: Experience setting up and maintaining modern CI/CD pipelines.
  • Observability: Practical experience implementing and managing monitoring and logging tools.
  • Networking: Solid understanding of TCP/IP, load balancing, DNS, and cloud-native networking within Kubernetes.
  • Operating Systems: Strong command-line skills and experience with Linux systems.

Soft Skills & Team Fit

  • High Ownership: Demonstrated ability to own a problem end-to-end, from investigation to resolution and preventative measures.
  • Small Team Mentality: Happy to be a generalist and switch context quickly between support tickets, operational toil reduction, and long-term project work.
  • Adaptability: A proven track record of rapidly learning and applying new technologies and tools. Equivalent experience with other clouds (AWS/Azure) or similar tools is highly valued.
  • Communication: Excellent verbal and written communication skills for documentation and interacting with non-technical stakeholders.

Bonus Points For

  • Familiarity with Service Mesh technologies (e.g., Istio).
  • Experience in security best practices within cloud and container environments (e.g., hardening, secrets management).
  • Certifications in GCP or Kubernetes (e.g., CKAD, CKA, Professional Cloud DevOps Engineer).

Company

Brevanhoward
London (82), United Kingdom

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Brevanhoward's careers site·first seen 10 Jul 2026·last verified 9 Sept 2026·How we source jobs

Similar jobs

  • Site Reliability Engineer / Senior Engineer at Deutsche BankLondon 10 Upper Bank Street, United Kingdom–match not yet calculated
  • Production Engineer at BarclaysGlasgow Campus, United Kingdom–match not yet calculated
  • Manager, Site Reliability Engineering at MastercardHarrogate, United Kingdom–match not yet calculated
  • Senior Caching SRE at BarclaysKnutsford, United Kingdom–match not yet calculated
  • Reliability Engineer at HelloFreshDerby, United Kingdom–match not yet calculated

Browse more jobs

  • Site Reliability Engineer jobs in United Kingdom
  • Platform Engineer jobs in United Kingdom
  • Systems Engineer jobs in United Kingdom
  • DevOps Engineer jobs in United Kingdom
  • Site Reliability Engineer jobs in United States
  • Site Reliability Engineer jobs in India