Senior Site Reliability Engineer
CISCO Systems
As part of Cisco’s Webex Engineering Group in London, you will own the design, delivery, and governance of Kubernetes-based microservices configurations, enabling scalable, declarative deployments. You’ll drive a GitOps-enabled platform with Argo CD, Helm/Overlays, and secure secrets, while advancing automation and observability to boost reliability. You’ll work across Dev, Staging, and Production to reduce risk and accelerate delivery, partnering with cross-functional teams to support hybrid collaboration at scale.
Responsibilities- Own the end-to-end provisioning and configuration of Kubernetes-based microservices (Deployments, Services, Ingress, ConfigMaps) using Kubernetes and Helm
- Build reusable base configs with environment overlays and manage safe promotion across dev, staging, production
- Drive GitOps practices using Argo CD and implement progressive delivery including canary releases and validation gates
- Codify autoscaling, resiliency, and secure secret management with HashiCorp Vault and Kubernetes Secrets
- Apply AI-assisted approaches to configuration-as-code to speed up development and reduce deployment risk
- Ensure governance and reproducibility with Open Policy Agent, including backup, restore, and disaster recovery
- Develop a self-service platform using Backstage and Argo CD to enable teams to deploy via declarative configs
- Focus on observability and performance improvements using Prometheus and Grafana
- Address complex configuration, policy, and data service challenges with a data-driven approach
- Collaborate with cross-functional teams to improve reliability and system maturity
- 8+ years in software development
- Strong Kubernetes and microservices experience
- Experience with GitOps (CI/CD)
- Analytical degree or equivalent experience
- Adaptable and problem-solver
- Analytical mindset
- Collaborative team player
- Kubernetes
- Microservices
- GitOps
Reference: WJ-747_30497001