IT & Software

Lead Cloud Site Reliability Engineer

Lloyds Banking Group

Manchester · Greater Manchester · United Kingdom

Overview

Lead SRE on the Public Cloud Platform, guiding a team to improve reliability and observability across GCP and Azure. You’ll partner with product and engineering leads to embed reliability into roadmaps and deliver scalable, observable cloud services. The role combines hands-on engineering with team leadership to foster learning, automation, and continuous improvement. A strong hook is shaping a resilient cloud platform that underpins the bank’s fintech ambitions.

Pay / Benefits
  • Competitive salary and performance-related bonus
  • 28 days holiday plus bank holidays
  • Generous pension contribution
  • Private medical insurance
  • Flexible benefits to suit your lifestyle
  • Hybrid working model and family-friendly policies
Responsibilities
  • Lead and coach a high-performing SRE team, fostering autonomy and inclusion
  • Embed reliability into roadmaps, backlogs and delivery decisions with Product Owners and Engineering Leads
  • Apply SRE concepts (SLIs, SLOs, error budgets) to ensure reliability and performance
  • Drive observability improvements across metrics, logs, traces and events
  • Utilise Dynatrace for dashboards and alerting
  • Own Infrastructure-as-Code and CI/CD environments, implement enhancements
  • Coordinate incident response and root cause analysis, lead post-incident reviews
  • Collaborate with multidisciplinary teams to reduce toil and improve operability
  • Contribute hands-on engineering as needed to validate decisions and promote best practice
  • Foster curiosity, experimentation and first-principles thinking to evolve engineering culture
Key requirements
  • Proven SRE experience in Azure, GCP, or both
  • Strong knowledge of SLIs, SLOs, and error budgets
  • Experience ensuring production reliability (availability, performance, recoverability)
  • Hands-on or leadership experience in incident and problem management
  • Background in software or cloud engineering with modern SDLC understanding
  • Practical experience with DevOps, CI/CD and automation
  • Experience improving observability in distributed systems
  • Ability to use data to prioritise reliability vs feature delivery
  • Excellent collaboration and communication with product, engineering and platform teams
  • Experience mentoring engineers and promoting inclusive culture
  • Collaborative/team-oriented
  • Mentoring and coaching
  • Curiosity and willingness to experiment
  • Azure and/or Google Cloud Platform (GCP) experience
  • SLIs/SLOs and error budgets
  • Observability tooling and strategies

Reference: WJ-747_30177227

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.