Lead Linux System Administrator
Morgan Stanley
In this Lead Infrastructure Production Management & Reliability Engineering role, you own the reliability and evolution of Morgan Stanley’s Linux-based infrastructure. You will work across a complex mix of open-source, in-house components, and critical Linux services to deliver a stable, secure platform with enhanced automation. Collaborating with global teams, you’ll drive incident prevention, capacity planning, and platform improvements that scale with business needs. This is a high-impact opportunity to shape operations for enterprise-grade systems in a dynamic financial-services environment.
Pay / Benefits- flexible working arrangements
- diverse and inclusive culture
- career growth opportunities
- global collaboration
- competitive employee benefits
- opportunities for internal mobility
- Support, deploy, and troubleshoot large-scale production Linux environments with emphasis on stability, performance, and reliability
- Contribute as an Infrastructure SRE with hands-on coding and automation
- Participate in global follow-the-sun on-call rotation for production issues
- Collaborate with engineers, developers, and project teams to implement infrastructure improvements
- Create and maintain clear documentation for systems, processes, and support models
- Adhere to change, incident, and problem management processes
- Engage in global forums to review projects, standards, and delivery priorities
- 5-7 years of enterprise-level experience with large distributed Linux environments
- Strong Linux knowledge, hardware platforms, automated builds, and lights-out management
- Ability to build from scratch, update configurations, and modify code paths
- Understanding of kernel bypass and latency-sensitive devices
- Solid TCP/IP networking knowledge, security fundamentals, implementation, and troubleshooting
- Experience with monitoring, alerting, and operational support solutions
- Analytical and troubleshooting skills in a fast-paced enterprise environment
- Coding or scripting experience, preferably Python
- Experience with CI, source control, release management, and change control
- Solid understanding of Linux kernel internals and performance tuning
- Experience with trading system configuration, deployment, and management
- Experience improving operational processes through automation, tooling, and documentation
- Collaborative cross-functional communication
- Prioritization and ability to navigate ambiguity
- Owner mindset and self-starter attitude
- Python scripting
- Linux system administration and automation
- Automated builds and lights-out management
Reference: WJ-747_30160792