Senior Site Reliability Engineer
Bango plc
Bango enables content providers to reach more paying customers through global partnerships. Bango revolutionized the monetization of digital content and services, by opening-up online payments to mobile phone users worldwide.
Today, the Digital Vending Machine® is driving the rapid growth of the subscriptions economy, powering choice and control for subscribers.
The world’s largest content providers, including Amazon, Google and Microsoft trust Bango technology to reach subscribers everywhere.
Bango, where people subscribe.
Role
As a Senior Site Reliability Engineer at Bango, you are a technical leader for the reliability, performance and continuous improvement of the Bango Platform. You hold everything expected of a Site Reliability Engineer — owning reliability end-to-end across the infrastructure, delivery pipelines, observability and incident response that keep the platform running to agreed service levels — but at greater scope, complexity and influence, and you take responsibility for lifting the capability of the engineering function around you.
Like the SRE role, you combine three things that have historically sat in separate teams: platform and cloud infrastructure engineering, automation and delivery pipeline ownership, and proactive/reactive reliability engineering including incident response and customer impact management. What sets the senior role apart is not simply doing this work to a higher standard, but becoming the technical centre of gravity for it — the person the team orients around on hard problems, who leads significant cross-cutting work end-to-end, and who is a steadying, trusted presence in the most serious incidents and the most contentious design decisions.
You take deliberate responsibility for the strength and depth of the SRE function, not just your own output — spotting capability gaps and key-person risks before they bite, growing engineers into harder work, spreading concentrated knowledge, and helping shape how SREs are hired, onboarded and developed. Because the role spans the third-line support team and the PICE team, you think about the bench across both. You do this as a senior individual contributor and technical leader, not a line manager: you lead through credibility, clear framing and judgment rather than positional authority or ownership of headcount.
At the furthest reach of the role, you help shape where Bango’s reliability and platform capability is heading — anticipating scaling limits, emerging risks, and investments that are cheap now and expensive later, and making grounded, evidenced cases for getting ahead of them. This forward-looking contribution is a stretch dimension of the role rather than a baseline expectation: a Senior SRE who is excellent across everything else is a strong, complete Senior SRE.
You are a senior member of the Managed Services & Support team, working alongside NOC, Partner Support and TSM colleagues, and collaborating closely with Software Engineering, Integration Engineering and InfoSec. You represent the reliability and platform perspective in technical discussions across the business — and, increasingly, that perspective shapes decisions before it is asked for.
Responsibilities
A Senior SRE carries all the responsibilities of the SRE role at a raised bar — broader scope, cross-team impact, and setting the standards others follow rather than meeting them within a single domain — together with the leadership responsibilities set out below.
Reliability & Incident Management
- Own the reliability of the Bango Platform across systems, leading the hardest and most cross-cutting investigations through to permanent, structural fixes that remove whole classes of recurring issues.
- Act as a calm, trusted incident lead in major or ambiguous incidents — bringing order, coordinating across teams, and owning customer-impact communication and post-incident learning.
- Drive proactive reliability improvement, anticipating and eliminating risk before it reaches customers rather than only responding to live issues.
- Work with Product and Integration Engineering to get recurring and systemic reliability issues onto roadmaps and permanently resolved.
Observability, Monitoring & Security
- Set the standard for observability across teams — raising signal quality, driving down alert fatigue, and introducing the SLI/SLO maturity that others adopt.
- Make poorly-understood systems legible, and design the monitoring patterns and tooling other engineers build on.
- Shape platform security posture and operational risk management with InfoSec — anticipating weaknesses before they are exploited or audited, and designing secure-by-default patterns other teams adopt.
- Own secrets, certificate and access-management practice within scope, and raise the bar for how the platform manages operational risk.
Platform & Infrastructure Engineering
- Design, build and operate platform services and cloud infrastructure at the hardest end, setting the reusable tooling, templates and integration patterns other SREs and teams adopt.
- Reduce duplication and fragmentation across the platform, and steer teams away from untracked workarounds toward supportable, standard solutions.
- Partner with the core platform team as a trusted, evidence-led voice — influencing their roadmap and architecture without overstepping their ownership.
- Ensure platform services scale efficiently and cost-effectively, treating cost and efficiency as part of engineering quality.
Automation & Software Delivery
- Own and continuously raise the standard of infrastructure-as-code, CI/CD and GitOps practice across the software delivery lifecycle.
- Eliminate whole categories of toil and technical debt rather than individual instances, and make the resulting pipelines and patterns the reference others copy.
- Drive adoption of automation tooling and good delivery practice across Bango engineering, not just within your own work.
- Work with Software Engineering on the automation and setup of new projects, and with QA on effective test automation.
Collaboration, Standards & Documentation
- Represent the reliability and platform perspective across the business so that it shapes decisions early — getting operability and supportability designed in by default.
- Communicate confidently at all levels, adjusting style appropriately for CEO, Chair, engineers or the People team.
- Define, document and steward the engineering standards, platform patterns and operational procedures the function works to.
- Create clear documentation and runbooks, participate in code review with constructive, detailed feedback, and use both to visibly raise the capability of the engineers around you.
Technical Leadership & Direction
- Set technical direction for reliability and platform work by framing problems clearly enough that the team adopts your approach because it is the best one, not because of title.
- Lead significant or cross-cutting work end-to-end, holding the shape of something complex while others contribute.
- Be the steadying technical presence in major incidents and contentious design decisions — the person the team looks to for order.
- Make the people around you more effective and aligned, deliberately avoiding becoming a bottleneck, and respecting the boundaries of the core platform team and Tech Leads.
Capability & Bench Building
- Take responsibility for the depth and resilience of the SRE function — spotting capability gaps and key-person risks across the third-line and PICE teams and working to close them.
- Grow engineers into harder work, spread concentrated knowledge, and deliberately develop successors for the critical roles, including behind yourself.
- Contribute meaningfully to how SREs are hired — articulating what 'good' looks like and holding the bar — and to how they are onboarded and developed.
- Strengthen the team as a system so it is more capable tomorrow than today and less fragile to any one person leaving. As a senior IC, this is capability and technical leadership, not ownership of headcount or formal people decisions.
Strategic & Forward-Looking Contribution
- Anticipate where Bango's reliability and platform needs are heading — scaling limits, reliability risks building with growth, investments cheap now and costly later — and shape direction before the need becomes urgent.
- Make grounded, evidenced cases that point at a specific, plausible future need paired with a pragmatic first step, distinguishing the shift worth getting ahead of from the shiny distraction.
- Bring a reliability-and-platform lens to direction-setting, informing and shaping strategy without overstepping into owning company technology strategy, which sits with the core platform team and engineering leadership.
- Where the evidence does not support a fashionable or premature bet, argue against it as readily as you argue for the investments that matter.
Essentials
- 5+ years' experience in a Cloud, Platform, DevOps or SRE role in a commercial environment, including demonstrable experience operating at a senior or lead level.
- A track record of technical leadership — setting direction others follow, leading significant cross-cutting work end-to-end, and being the person peers bring hard, ambiguous problems to.
- Experience leading major incidents: coordinating across teams under pressure, owning customer impact, and driving structural resolution.
- Demonstrable impact growing other engineers through mentoring and code review — raising the capability of a team, not just delivering individually.
- Deep production experience with a major cloud provider (AWS, Azure or GCP).
- Strong Linux administration and troubleshooting (process management, memory management, signals, etc.).
- Production experience with containerisation and orchestration (Docker, Kubernetes).
- Strong Infrastructure-as-Code experience (Terraform or equivalent), with a track record of building reusable patterns others adopt.
- CI/CD and GitOps experience (GitLab CI/CD or equivalent), with a clear, well-formed view of what 'good' continuous delivery looks like and how to drive it across teams.
- Strong scripting ability in Python and/or Bash.
- Solid networking fundamentals — TCP/IP, DNS, HTTP, TLS.
- Experience with relational and non-relational databases (e.g. SQL Server, MySQL, DynamoDB, Elasticsearch).
- Strong incident management, monitoring and root cause analysis practice, including designing the observability others rely on.
- Excellent troubleshooting and analytical problem-solving, individually and leading others.
- Working experience in Agile teams and development processes.
- Excellent oral and written communication with technical and non-technical audiences, up to and including executive level.
- Calm, clear-thinking and flexible under time pressure when resolving live issues and leading others through them.
- Strong grasp of low- and high-level architectural concepts and common cloud design patterns.
Desirables
- Experience shaping hiring bars, onboarding, or capability frameworks for an engineering function.
- Experience contributing a reliability or platform lens to technical strategy and roadmap decisions.
- Configuration management tooling (Puppet, Ansible).
- Advanced Kubernetes (Karpenter, KEDA, HPA/VPA, Service Mesh).
- Observability tooling (Datadog, OpenTelemetry) and mature SLI/SLO design.
- Security and identity tooling (SSO, IAM, PKI).
- Queue or streaming design patterns (SQS, Kafka).
- Resilience and disaster-recovery design, including multi-region or chaos/resilience testing.
- Commercial or low-level languages beyond scripting (Java, C#, Golang, C++, Rust).
- Experience with legacy Windows Server / .NET / T-SQL / PowerShell environments, where Bango still operates them.
- Experience working across time zones and cultures on cross-border projects.
- Experience with virtualisation, SAN, or data security and protection practices.
Benefits
- A friendly, informal working environment
- Your own Bango buddy – to help you settle in
- Bendi-time (flexible working hours)
- Bango social events
- Choose your own headphones, keyboard & mouse
- Generous share option scheme
- Private Medical Insurance
- Health Cash Plan
- 25 days holiday a year increasing to 28 days with 4 years’ service
- Cycle to work, gym discount
- Weekly Pilates & Yoga classes (virtual)
- Financial support for employee activity groups and charitable activities
- Free fruit, drinks and snacks, limitless tea, coffee and good quality espressos
- Company branded hoodie… to keep you happy and comfortable
- Group personal pension scheme
- Life assurance
- Employee Assistance Program
- 1Password
- Income Protection
- Bango branded Chilly’s bottle and coffee cup
Please read our Privacy Policy below before proceeding to Application
Privacy Policy.pdf
Interested in this exciting opportunity?
#J-18808-LjbffrReference: WJ-766_22291896