openqareer

Site Reliability Engineer

Bango.net Limited · Cambridge

# Site Reliability Engineer **Bango.net Limited** · Cambridge · `On-site` 🕒 **Статус:** *Опубликовано: вчера* · *Источник: Indeed* --- ### About the Role Bango enables content providers to reach more paying customers through global partnerships. Bango revolutionized the monetization of digital content and services, by opening-up online payments to mobile phone users worldwide. Today, the Digital Vending Machine® is driving the rapid growth of the subscriptions economy, powering choice and control for subscribers. The world’s largest content providers, including Amazon, Google and Microsoft trust Bango technology to reach subscribers everywhere. Bango, where people subscribe. Role As a Site Reliability Engineer at Bango, you own the reliability, performance and continuous improvement of the Bango Platform end-to-end — from the infrastructure and pipelines that build and deploy it, to the observability and incident response that keep it running to agreed service levels. You combine three things that have historically sat in separate teams: platform and cloud infrastructure engineering, automation and delivery pipeline ownership, and proactive/reactive reliability engineering including incident response and customer impact management. You design, build and operate the automation, observability and platform capabilities that other engineering teams rely on, and you are equally comfortable diagnosing a live incident as you are designing the Terraform module that prevents the next one. You work closely with Product and Integration Engineering teams to understand what ‘good’ looks like, support the delivery of well-architected, secure and observable services, and represent the reliability and platform perspective in technical discussions. You are a key contributor to Bango’s engineering standards and practices, and you take pragmatic, evidence-based decisions about where to invest effort — whether that’s automating a manual process, hardening a security control, or resolving a live customer-impacting issue. You are a member of the Managed Services & Support team, working alongside NOC, Partner Support and TSM colleagues, and collaborating regularly with Software Engineering, Integration Engineering and InfoSec. Responsibilities Reliability & Incident Management Ensure the Bango Platform is available, responsive and performant for customers; work proactively to identify and mitigate risk, and reactively to resolve active incidents. Investigate incidents, perform root cause analysis, and drive permanent corrective action rather than workarounds. Understand Bango's customers, their expectations, and the business impact of incidents on the platform or with partners. Liaise frequently with Product and Integration Engineering to document issues and defects, describe their impact, and support timely resolution — including recurring minor issues that risk being overlooked. Support the detection of issues introduced by new deployments, and are empowered to roll these back or work around them. Help prioritise bug fixes in collaboration with product teams, bringing ideas and potential solutions to the table. Observability, Monitoring & Security Design and maintain observability, monitoring and alerting capabilities that give all Bango teams visibility of service health and availability. Identify and remediate false alarms or gaps in current monitoring, continuously optimising signal quality. Implement and maintain platform security controls, hardening standards and access management, including secrets and certificate lifecycle management. Continuously improve platform security posture and operational risk management in collaboration with InfoSec. Platform & Infrastructure Engineering Design, implement, operate and continuously improve platform services and cloud infrastructure capabilities. Take ownership of platform services throughout their lifecycle, including cloud infrastructure, Kubernetes platforms and supporting tooling. Develop and maintain reusable platform components, templates and engineering standards that other teams can adopt. Ensure Cloud configuration and automation software aligns to Bango standards, doesn't needlessly re-invent existing solutions, and enables other teams to integrate effectively. Support the evolution of Bango's cloud and hybrid architecture. Continuously optimise infrastructure cost and ensure platform services are designed to scale efficiently with demand. Automation & Software Delivery Design, build and maintain infrastructure-as-code, CI/CD and GitOps workflows across Bango's software delivery lifecycle. Own and continuously improve build and deployment pipelines, ensuring reliable, repeatable releases that meet automated quality and security gates. Identify and eliminate manual activity, automation bottlenecks and technical debt across delivery and operational processes. Advocate for and drive adoption of automation tooling and practices across Bango engineering. Work with Software Engineering on the automation and setup of new projects, and with QA on effective test automation. Collaboration, Standards & Documentation Act as the reliability and platform representative for cross-functional teams — Developers, QA, Product Managers and Technical Integration Specialists — with stakeholders across the business. Confidently communicate at all levels, adjusting style appropriately for CEO, Chair, engineers or the People team. Define, document and promote engineering standards, platform patterns and operational procedures. Create clear documentation and runbooks to guide Support teams and help Engineering resolve issues permanently. Support onboarding, mentoring and knowledge-sharing, and act as a key contributor to Bango's engineering principles and best practices. Participate in code review and provide constructive, detailed feedback. Call out areas where Bango can improve service or reduce cost. Essentials 3+ years' experience in a Cloud, Platform, DevOps or SRE role in a commercial environment. Production experience with a major cloud provider (AWS, Azure or GCP). Strong Linux administration and troubleshooting (process management, memory management, signals, etc.). Production experience with containerisation and orchestration (Docker, Kubernetes). Infrastructure-as-Code experience (Terraform or equivalent). CI/CD and GitOps experience (GitLab CI/CD or equivalent), with a solid understanding of what 'good' continuous delivery looks like. Strong scripting ability in Python and/or Bash. Solid networking fundamentals — TCP/IP, DNS, HTTP, TLS. Experience with relational and non-relational databases (e.g. SQL Server, MySQL, DynamoDB, Elasticsearch). Practical experience of incident management, monitoring and root cause analysis. Strong troubleshooting and analytical problem-solving, individually and as part of a team. Working experience in Agile teams and development processes. Excellent oral and written communication with both technical and non-technical audiences. Calm, clear-thinking and flexible under time pressure when resolving live issues. Understanding of lower- and high-level architectural concepts and common cloud design patterns. Desirables Configuration management tooling (Puppet, Ansible). Advanced Kubernetes (Karpenter, KEDA, HPA/VPA, Service Mesh). Observability tooling (Datadog, OpenTelemetry) and SLI/SLO design. Security and identity tooling (SSO, IAM, PKI). Queue or streaming design patterns (SQS, Kafka). Commercial or low-level languages beyond scripting (Java, C#, Golang, C++, Rust). Experience with legacy Windows Server / .NET / T-SQL / PowerShell environments, where Bango still operates them. Experience working across time zones and cultures on cross-border projects. Experience with virtualisation, SAN, or data security and protection practices. Benefits A friendly, informal working environment Your own Bango buddy – to help you settle in Bendi-time (flexible working hours) Bango social events Choose your own headphones, keyboard & mouse Generous share option scheme Private Medical Insurance Health Cash Plan 25 days holiday a year increasing to 28 days with 4 years’ service Cycle to work, gym discount Weekly Pilates & Yoga classes (virtual) Financial support for employee activity groups and charitable activities Free fruit, drinks and snacks, limitless tea, coffee and good quality espressos Company branded hoodie… to keep you happy and comfortable Group personal pension scheme Life assurance Employee Assistance Program 1Password Income Protection Bango branded Chilly’s bottle and coffee cup Interested in this exciting opportunity? We’d love to hear from you! Location UK - Cambridge Department MS&S Job Title Site Reliability Engineer

Наблюдалась 2026-09-15, впервые 2026-09-14, источник — Indeed.

Открыть у работодателя