Senior Site Reliability Engineer
CISCO Systems
In this role you help design and operate Kubernetes-based microservices for a modern collaboration platform. You work within the Webex Engineering team to deliver reliable, scalable configurations, using GitOps practices and secure, governed deployments. You will leverage AI-assisted tooling to accelerate configuration and reduce risk, while improving observability and performance in a hybrid-work environment. This is a chance to shape deployment pipelines, promote standard processes, and enable self-service for engineering teams, at scale.
Responsibilities- Own design, coding, testing, and delivery of Kubernetes-based microservices configurations (Deployments, Services, Ingress, ConfigMaps) using Kubernetes and Helm
- Build reusable base configs with environment overlays and enable safe promotion across dev, staging, production
- Drive operational excellence in a GitOps model using Argo CD and implement progressive delivery with canary releases and validation gates
- Codify autoscaling and resiliency, manage secure secrets via HashiCorp Vault and Kubernetes Secrets
- Use AI-assisted tools to accelerate configuration-as-code development, validation, and optimization
- Address complex challenges across configuration, policy, observability, and data services with a data-driven approach (Prometheus, Grafana)
- Own end-to-end configuration quality and governance with Open Policy Agent; ensure secure, compliant deployments and full reproducibility (backup, restore, DR)
- Advance a self-service platform using Backstage and Argo CD, enabling declarative deployments and standard processes
- 8+ years software development experience
- Strong Kubernetes and microservices experience
- Experience with GitOps (CI/CD)
- Analytical degree or equivalent experience
- Adaptable and problem-solving mindset
- cross-functional collaboration
- data-driven decision making
- Kubernetes
- microservices
- Helm
Reference: WJ-747_30137310