Software Engineer/ SRE (Linux)
Visa
Overview
In this role you will help Visa’s cloud platform remain reliable and scalable, enabling developers to focus on innovation. You will promote observability and automate issue resolution, collaborating with software teams to support security, availability, and performance. You’ll triage incidents, design robust systems, and maintain monitoring and CI/CD integrations. The role blends SRE discipline with DevTools expertise to improve developer productivity and platform reliability at scale.
Responsibilities- Primary DevTools support for GitHub, Jenkins, Jira, Artifactory; troubleshoot tool-related issues to minimize downtime
- Maintain and optimize CI/CD pipelines and integrations for reliability and scalability
- Collaborate with development teams to improve workflows and automation
- Design, implement, and maintain systems with high availability, scalability, and performance
- Monitor reliability and lead incident response and root cause analyses
- Develop and maintain observability solutions (metrics, logging, tracing)
- Participate in on-call rotations and advance automation and self-service.
- Document processes, troubleshooting guides, and reliability playbooks
- Bachelor's degree in IT, CS or related field or 3+ years of relevant work experience
- 0.5–3 years in SRE and/or DevTools support roles
- Proficiency in at least one DevTool (GitHub, Jenkins, ArgoCD, Jira, Artifactory)
- Basic programming/scripting in Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, CloudFormation
- Solid Linux and/or Windows experience; understanding of distributed computing environments
- Experience with CI/CD tooling (Jenkins, GitHub, Bitbucket, ArgoCD, Artifactory, Azure DevOps)
- Observability tooling experience (Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry)
- Experience with relational and non-relational databases (MySQL, MongoDB, PostgreSQL)
- Platform/SRE/Production Engineering background for high-availability environments
- Distributed container platforms management (deployment, capacity, workload management) and container-first transformations
- On-call support for 24/7 operations
- Cloud platform experience
- Programming-scripting: Python, Ansible or similar
- Problem-solving
- Systems thinking
- Self-starter
- DevTools proficiency (GitHub, Jenkins, ArgoCD, Jira, Artifactory)
- CI/CD principles and pipelines
- Linux systems and networking
Reference: WJ-747_30171848