IT & Software

Senior Site Reliability Engineer

Brevan Howard

London · Greater London · United Kingdom

Overview

As a Senior SRE, you will drive the reliability, scalability, and performance of our core platform on GCP, with strong autonomy and ownership. You’ll split time between providing expert operational support for critical systems and leading new infrastructure projects that shape our cloud-native stack. Expect a proactive, problem‑solving role focused on automation, IaC, and robust observability to prevent issues before they arise. This opportunity lets you influence architecture and practices in a small, agile team. You’ll work closely with cross‑functional teams to deliver reliable, scalable services that enable the business to move fast.

Responsibilities
  • Architect, deploy, and maintain scalable, reliable infrastructure on GCP using Kubernetes and IaC tools
  • Champion automation across the SDLC using IaC, Python, and Bash
  • Own declarative infrastructure with Terraform and Helm for cloud resources and Kubernetes deployments
  • Implement and manage monitoring, alerting, and logging for visibility and proactive issue detection
  • Define and enforce SLOs/SLIs; participate in on-call rotations and lead post-incident reviews
  • Collaborate with software teams to guide deployment strategies, scalability, and cloud-native best practices
  • Take full ownership of projects from inception through production operation, including documentation and knowledge transfer
Key requirements
  • 4+ years hands-on experience with Google Cloud Platform (GCP) or similar
  • Expert-level Kubernetes production experience
  • Deep Terraform expertise for cloud and Kubernetes resources
  • Strong Helm experience for packaging and deploying on Kubernetes
  • Proficient in at least one major programming language, preferably Python
  • Experience setting up and maintaining modern CI/CD pipelines
  • Experience with monitoring and logging tooling
  • Solid understanding of TCP/IP networking, load balancing, DNS, and cloud-native networking in Kubernetes
  • Linux command-line proficiency
  • Excellent communication for documentation and stakeholder interaction
  • High ownership
  • Small team mentality
  • Adaptability
  • GCP
  • Kubernetes
  • Terraform

Reference: WJ-747_30153428

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.