IT & Software

Service Reliability Engineer - Manchester

Fitch Ratings

Manchester · Greater Manchester · United Kingdom

Overview

In this role you embed with Fitch Ratings development squads in Manchester to deliver reliable, scalable services at scale. You will guide Kubernetes adoption, drive DevOps tooling, and own observability to reduce incidents. You’ll mentor engineers and partner with cross-functional teams to enhance AI-enabled operations and CI/CD practices. This is a hands-on SRE role focused on cloud-native reliability within a leading financial information firm.

Pay / Benefits
  • Hybrid work (2–3 days in office)
  • Dedicated trainings and mentorship
  • Retirement planning
  • Tuition reimbursement
  • Comprehensive healthcare
  • Parental leave
Responsibilities
  • Lead delivery of reliable, scalable, mission-critical services
  • Guide squads on Kubernetes and modern deployment patterns
  • Mentor associate engineers and establish best practices
  • Design and advance service builds, DevOps tooling, and operational excellence
  • Architect and govern GitHub Actions CI/CD with quality gates, canary/blue-green deployments, and AI-assisted redeploy checks
  • Own observability in Datadog: define SLIs/SLOs, dashboards, alerts, and integrations
  • Champion AI-enabled operations using AWS Bedrock/SageMaker and MCP for log analysis and incident triage
  • Define and enforce cloud guardrails and security controls with IAM, SCPs, OPA policies, and centralized logging
  • Influence cross-functional roadmaps, lead release planning, and drive platform initiatives across CI&PE; participate in L3 on-call rotation
Key requirements
  • Deep hands-on experience in SRE, DevOps, or Platform Engineering across AWS and Azure
  • Proven production experience with Docker and Kubernetes
  • Linux and Windows administration; IIS/.NET and Java Spring Boot apps
  • Built and maintained CI/CD pipelines (GitHub Actions; Bamboo a plus) with DevSecOps practices
  • Scripting in Python, PowerShell, or Bash
  • Cloud security practices (IAM, secrets management) and basic infrastructure skills (Networking, Storage, DNS)
  • APM/telemetry tooling knowledge
  • Collaboration and teamwork
  • Mentoring and coaching
  • Effective communication
  • AWS
  • Azure
  • Docker

Reference: WJ-747_30148332

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.