Lead Infrastructure Engineer - AWS Cloud Support Engineering
JP Morgan Chase
In this Lead Infrastructure Engineer role, you will shape the infrastructure engineering discipline within Cloud Foundational Services, driving scalable, resilient systems for a global financial services leader. You collaborate with cross-functional teams to modernize technology processes and resolve complex issues at scale. You’ll leverage AI-assisted tooling to improve code quality and delivery, while maintaining secure and compliant operations. This role offers impact across on-call support, automation, and product-facing improvements that advance the company’s technology platform.
Responsibilities- Apply technical expertise to projects of moderate scope
- Drive a workstream across multiple infrastructure technologies
- Collaborate with other platforms to implement changes and modernization
- Develop creative designs and solutions for moderate-complexity problems
- Advise on upstream/downstream data and system implications and mitigation actions
- Leverage enterprise-authorized AI coding assist tools to improve quality and speed
- Utilize SDLC tools and AI-assisted development to enhance automation value
- Manage stakeholder relationships within compliance and SLAs
- Partner with engineering and product teams to close issues and drive product improvement
- Provide follow-the-sun 24/7 on-call support via a distributed time-zone-spread team
- Formal training or certification on infrastructure engineering concepts
- Deep knowledge in areas such as hardware, networking, databases, storage, deployment, automation, scaling, resilience, or performance assessments
- Strong experience with AWS core services (compute, storage, networking, monitoring, database technologies) in production
- Deep knowledge of cloud infrastructure and multi-cloud capabilities
- Familiarity with scripting languages (e.g., Python) and cloud certifications
- Hands-on experience with AI-assisted software development tools and ability to evaluate AI outputs
- Understanding of responsible AI use, data sensitivity, and security/resiliency expectations
- Experience in large distributed systems across compute, databases, messaging, observability, and telemetry
- Knowledge of incident, change, and problem management processes
- Understanding of data-driven decision making
- Familiarity with large-scale cloud migration or modernization initiatives
- Exposure to observability and telemetry tooling in complex environments
- Preferred: Kubernetes, Azure, GCP, or Terraform certifications
- Experience with follow-the-sun on-call support
- Ability to learn and teach in feature-rich product environments
- Strong collaboration and stakeholder management
- Continuous learning and knowledge sharing
- Analytical thinking and problem solving under pressure
- AWS core services (compute, storage, networking, RDS, DynamoDB, Aurora)
- Cloud infrastructure across public/private clouds
- Scripting with Python
Reference: WJ-747_30160546