Software Engineer III, Site Reliability Engineering, GCE AI
In this SRE role at Google Cloud you will build and maintain scalable, fault-tolerant systems to meet customer needs. You will work closely with cross-functional teams to improve reliability, capacity, and performance. You will review code, triage issues, and contribute to documentation, while participating in design discussions. The role emphasizes automation, infrastructure, and scalable software solutions in a collaborative, risk-tolerant environment. You will contribute to shaping large-scale systems and drive measurable improvements in uptime and efficiency.
Responsibilities- Write product or system development code
- Review code and provide feedback for best practices
- Contribute to documentation and education content
- Triage and debug/track issues affecting hardware, network, or services
- Participate in or lead design reviews to decide among technologies
- Bachelor’s degree in Computer Science, a related field, or equivalent practical experience
- 2 years of software development experience in one or more programming languages
- 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems
- collaboration
- problem solving
- analytical thinking
- software development
- design, analysis, and troubleshooting of distributed systems
- automation and infrastructure
Reference: WJ-747_30184199