Site Reliability Engineer , Cryptography, Access and Identity Services
AmazonWebServices
Overview
In this role you build and operate cloud services at massive scale, focusing on automation and operational excellence. You will mentor junior teammates and own cryptography, access, and identity services in a high-availability environment. You’ll use Linux and networking know-how to troubleshoot issues, improve capacity metrics, and reduce manual toil with AI-assisted tooling. You’ll work with cross-functional teams to shape resilient, scalable systems that power AWS customers.
Pay / Benefits- Mentorship & Career Growth
- Work/Life Balance
- Diverse experiences
- Inclusive culture
- Build, deploy and run services for customers at scale
- Automate service operations, deployment methods, and manual tasks
- Mentor and guide junior engineers
- Troubleshoot Linux-based systems and network connectivity issues
- Analyze trends/metrics to identify improvement opportunities
- Develop and refine operational procedures and tooling
- Maintain security and reliability in a 24/7 production environment
- Collaborate with teams to reduce toil and increase efficiency
- Investigate root causes for customer issues and implement durable fixes
- Experience in site reliability engineering (SRE) or related fields
- Experience working with Linux
- Experience in systems engineering
- Proficiency in Python, Java, Perl, PHP, Ruby, Bash, Shell or equivalent
- Leadership
- Mentorship
- Problem solving
- Linux
- Python/Java/ scripting languages
- TCP/IP and networking protocols (HTTP, DNS)
Reference: WJ-747_30154261