Senior Site Reliability Engineer
London Stock Exchange Group
In this Senior SRE role you will shape the reliability foundations across Risk Intelligence platforms and partner with Architecture, Engineering, Security, and Platform teams to embed reliability from day one. You will lead observability implementations, drive incident reduction, and ensure security and compliance alignment. The role offers impact through designing resilient patterns, guiding cost-efficient cloud strategies, and mentoring engineers in a collaborative, high‑security environment.
Pay / Benefits- healthcare
- retirement planning
- paid volunteering days
- wellbeing initiatives
- tailored benefits and support
- Establish SRE foundations for new projects, including environments, monitoring, alerting, and operational readiness
- Define and champion observability standards across metrics, logs, traces, and SLIs/SLOs
- Design and evolve monitoring/alerting to improve visibility and reduce toil
- Drive reliability improvements via incident reduction, performance tuning, and resilient patterns
- Collaborate with Security to meet compliance, security, and risk-management expectations
- Influence architecture with data-driven cloud cost optimization and efficiency initiatives
- Mentor engineers, shaping engineering standards and fostering learning
- Work with cross-functional teams to ensure reliability is built into system design from the start
- 5+ years hands-on experience in SRE, Platform Engineering, Infrastructure, or related roles
- Strong experience with AWS (or Azure), including EKS, ECS, EC2, networking, IAM, and managed services
- Solid understanding of cloud security principles and collaboration with security teams
- Strong Linux systems administration background
- Proven experience designing and operating observability platforms (monitoring, logging, alerting)
- Hands-on experience with Datadog for metrics, logs, APM, and alerting
- Strong understanding of SRE principles (SLIs/SLOs, error budgets, incident management)
- Experience collaborating with architecture and engineering teams on system design and delivery
- Experience with cloud cost optimization strategies and tooling
- Strong collaborative mindset
- Leadership presence and ownership
- Calm and methodical under pressure
- Datadog (metrics, logs, APM, alerts)
- AWS services: EKS, ECS, EC2, IAM, networking
- Linux system administration
Reference: WJ-747_30160012