Senior Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Senior Site Reliability Engineer roles and what employers typically expect.
Senior Site Reliability Engineer Meaningful technical ownership Broader engineering influence Solve reliability problems properly rather than work around them The best Site Reliability Engineers do more than keep systems running. They improve how software is designed. They remove recurring sources of operational pain. They build tools that make engineering teams faster and safer. They see risks before they become incidents and leave systems easier to understand, operate and change than they found them. That is the kind of Senior SRE we are looking for at Pushpay. This is an opportunity to work across software, cloud infrastructure, reliability, security and engineering practices. You will take on complex problems that do not always arrive neatly defined, follow them across system boundaries and turn them into durable improvements. You will have the space to think deeply, influence architecture and raise the operational capability of the teams around you. The work You will work on the parts of engineering that become most interesting as systems and organisations grow. That may mean investigating a difficult production issue that crosses application, infrastructure and data boundaries. It may mean identifying a pattern behind repeated alerts and engineering the underlying problem away. It may mean building automation that removes manual work, improving the operational visibility of a service or challenging a design before it creates reliability problems in production. You will contribute to operability and architecture reviews with engineering teams, helping shape systems that are reliable, secure and practical to run—not just technically elegant on paper. This is a role for someone who enjoys working with code as much as infrastructure and who sees reliability as an engineering discipline rather than a collection of operational tasks. As a senior member of the team, you will be trusted to: Take end-to-end ownership of complex reliability and systems work Recognise operational, architectural and security risks early Turn broad or ambiguous problems into clear, well-scoped improvements Build tools and automation that reduce toil and improve engineering velocity Improve the observability, operability and resilience of production systems Help teams make stronger technical decisions before software reaches production Communicate trade-offs, risks and progress clearly Mentor engineers who are developing their systems and operational knowledge Contribute to technical hiring and help strengthen the wider engineering team You will participate in on-call support when rostered, but the objective is not simply to respond well when something goes wrong. It is to learn from operational signals, remove recurring failure modes and steadily improve the systems and practices behind them. You will work with Software Engineers, QA, Product, Delivery, Engineering Managers and other parts of the business. You will help teams understand how design choices affect reliability, security, supportability and customer outcomes. You will be encouraged to question existing approaches, refine engineering practices and contribute to the shared standards and knowledge that shape how Pushpay builds and operates software. What you’ll bring We are looking for someone with strong experience across site reliability engineering, systems engineering or software engineering, ideally gained while developing or operating multi-user web, mobile or cloud software products. Your experience may include several of the following: Cloud platforms such as AWS, GCP or Azure Infrastructure as Code using Terraform or CloudFormation Python, Go, Rust, JavaScript or .NET Command line environments such as Bash and PowerShell Relational or document-based databases HTTP, SSL/TLS, REST APIs or GraphQL Git and distributed version control GitHub Actions, Jenkins or other CI/CD tooling You will need to be able to reason across an end-to-end system, communicate clearly in writing and convers…