IT & Software

Senior Service Reliability Engineer

Barclays

Knutsford · Cheshire · United Kingdom

Overview

In this role you will lead IT Services with a focus on reliability, risk management, and governance within a major banking environment. You will coordinate across engineering, infrastructure, security and product teams to maintain stable, secure platform services and manage incidents and changes. You’ll drive improvements, oversee releases, and ensure alignment with regulatory and internal controls. This position offers a chance to shape platform reliability at scale and influence senior stakeholders.

Responsibilities
  • Develop and implement strategic direction for IT Services and modern governance practices
  • Oversee IT Services department performance, objectives, and efficiency
  • Stakeholder relationship management and external third-party service oversight
  • Policy and procedure development, control targets, SLA adherence and incident/problem/change governance
  • Identify and mitigate IT Services risk; align change and compliance functions
  • Monitor financial performance and optimize costs and revenue within IT Services
  • Direct IT Services projects, including research, product launches, and integrated client solutions
  • Maintain critical technology infrastructure, lead complex technical issue resolution with minimal operational disruption
  • Coordinate platform releases, drive operational improvements, and monitor service reliability
  • Collaborate with Engineering, Infrastructure, Security, and Product teams to improve monitoring, automation, and resilience
  • Provide clear stakeholder communication during service events and governance activities
Key requirements
  • Extensive leadership in Production Support, SRE, and Platform Operations
  • Strong ITIL, Major Incident Management, problem management, and operational resilience experience
  • Hands-on AWS cloud technologies and container platforms (OpenShift/Kubernetes)
  • Experience with observability tooling, automation, CI/CD, Infrastructure as Code
  • Proven ability to lead incidents, perform root cause analysis, and communicate with senior stakeholders
  • Knowledge of operational risk, governance, regulatory compliance, and cloud security in regulated environments
  • Platform ownership with KPI/SLA management, capacity planning, and continuous service improvement
  • Strategic leadership and ability to influence enterprise stakeholders
  • leadership and people development
  • cross-functional collaboration
  • effective communication with senior stakeholders
  • AWS
  • OpenShift
  • Kubernetes

Reference: WJ-747_30875220

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.