IT & Software

Staff Technical Program Manager, Site Reliability Engineering

MongoDB

New York · New York · United Kingdom

Overview

As a TPM for SRE, you partner with SRE leaders to scale MongoDB’s cloud platform, driving program execution and reliability practices. You coordinate across US and EMEA teams to deliver smoother launches, clearer roadmaps, and stronger reliability metrics. You’ll shape scalable processes and reduce hero culture by empowering teams to operate independently. This role offers remote work from the East Coast and the chance to impact platform reliability at scale.

Pay / Benefits
  • remote work options
  • equity and employee stock purchase program
  • flexible paid time off
  • 20 weeks fully-paid gender-neutral parental leave
  • fertility and adoption assistance
  • health benefits including mental health support
Responsibilities
  • Define and drive program planning and execution with SRE engineers and leaders; manage dependencies and track work in Jira to ensure on-time delivery
  • Lead production reliability efforts, including change management, launch readiness, and defining operational SLOs/SLIs
  • Coordinate cross-functional efforts with Security, Compliance, Cloud platform, and other engineering teams; drive incident response and follow-through
  • Design lightweight frameworks and processes to enable scalable, reliable delivery and reduce reliance on individual heroes
Key requirements
  • 8+ years in technical program management, engineering management, or similar role with software engineering teams
  • Proven track record leading large-scale, cross-team platform initiatives through ambiguity and change
  • Strong knowledge of production change management, SDLC, and reliability metrics (SLOs/SLIs)
  • Skilled at shaping roadmaps and managing dependencies
  • Ability to query and interpret metrics, logs, or data sources to inform decisions and communicate risk
  • Excellent communicator—clear, concise, calm—across engineers, partners, and executives
  • Low-ego, highly collaborative, ownership mindset for end-to-end problems
  • collaboration
  • clear communication
  • ownership of hard problems
  • Kubernetes (Nice to Have)
  • cloud networking (Nice to Have)
  • observability stacks (metrics, logs, tracing, alerting)

Reference: WJ-747_30149279

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.