IT & Software

Senior Platform Reliability Engineer

Ricoh

London · Greater London · United Kingdom

Overview

Senior Platform Reliability Engineer at Ricoh in London. You will own the reliability, resilience and operational integrity of hybrid and cloud platforms, delivering high availability and secure, compliant services. You’ll combine hands-on engineering with automation, IaC and observability to reduce toil and improve recovery. You’ll work closely with SRE and helpdesk teams to keep services meeting KPIs, while driving automation-led improvements. This is a chance to help shape scalable, secure infrastructure within a values-driven, collaborative environment.

Pay / Benefits
  • fair rewards
  • flexible working
  • wellbeing resources
  • recognition programmes
  • volunteering and community programmes
  • inclusive and supportive culture
Responsibilities
  • Define and deliver availability, latency, performance, capacity and scalability targets
  • Lead root-cause analysis and incident/problem management for major incidents
  • Promote blameless post-mortems with tracked actions
  • Drive infrastructure-as-code and automation across Azure and co-located environments
  • Evolve image creation pipelines for secure, repeatable server images
  • Embed observability with metrics, logs, traces and alerting
  • Collaborate with SRE and helpdesk teams to deliver reliable services
  • Oversee automated patching, vulnerability remediation and configuration compliance
  • Introduce KPIs and dashboards for reliability, incident trends, MTTR, change failure rate and capacity
  • Travel occasionally to data centres and offices to ensure alignment with business requirements
Key requirements
  • Strong business awareness with a clear link between infrastructure services and business operations
  • Proven experience in a similar role with hands-on infrastructure, cloud and operations
  • Extensive Azure experience (IaaS, PaaS, networking, identity, storage) and on-prem data centre operations
  • Hands-on IaC and configuration management (Terraform, ARM/Bicep, Ansible, PowerShell DSC)
  • CI/CD pipelines experience (Azure DevOps, GitHub Actions)
  • Monitoring/observability expertise, alert design and dashboarding
  • Commercial awareness including vendor/partner engagement and cost optimisation
  • IT service management mindset with strong communication and documentation
  • ISO 27001 or similar regulated environment experience
  • ITSM integration (e.g., ServiceNow) and ITIL process alignment
  • Understanding of business/application impact of infrastructure decisions
  • Knowledge of security, architecture and vendor/commercial considerations
  • customer-focused mindset
  • clear communication
  • documentation
  • Azure (IaaS, PaaS, networking, identity, storage)
  • Terraform
  • ARM/Bicep

Reference: WJ-747_30137827

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.