Senior Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Senior Site Reliability Engineer roles and what employers typically expect.
**About Us** At [Félix](https://felixpago.com/), we're building the financial ecosystem for Latin immigrants in the U.S., starting with a revolution in remittances. Our core product is an AI-powered chatbot built on WhatsApp, allowing our users to send money home as easily as sending a text message. We leverage cutting-edge technology like AI, blockchain, and stablecoins to make cross-border payments faster, more affordable, and more accessible than ever before. We are a hyper-growth Series B company, backed by over $100 million in funding from top-tier global investors, including QED, Castle Island, Switch Ventures, HTwenty, Monashees, and General Catalyst Customer Value Fund. This isn't just about the numbers; it's a testament to the trust our investors have in our vision and our team. Additionally, Félix was selected as an “Endeavour Entrepreneur” and was a recipient of the CrossTech Fintech Startups Award. We are a group of extremely talented and dedicated high-performers, united by our shared obsession with a single goal: empowering our customers. We are all owners of Félix, driven by a bias for action and a true experimentation spirit to get shit done with urgency and focus. Joining Félix means you will be part of a team building a legacy, a company that will outlive us all. This is a rare opportunity to apply your skills to a deeply meaningful mission—serving a community that has been underserved for too long. We are a team that is fiercely loyal to each other, where radical transparency and constructive feedback are how we grow and push for excellence. We are bold, we care less about what others are doing, and more about creating sustainable value and a product that truly makes our users' lives better. We are building the future, today. **About the Role** As a Senior Site Reliability Engineer you will be a critical part of our Platform Engineering team. You will be responsible for creating automations to provide velocity to product engineering . This is a hands-on role for a builder who is passionate about shifting security left and empowering developers to ship secure code, quickly and confidently. You will be instrumental in maturing our DevSecOps practices, building out our automations. **Responsibilities** - Manage and optimize our infrastructure on Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE). - Automate provisioning and configuration using Terraform, Helm, and scripting languages such as Go, Python, and Bash. - Build, maintain, and improve monitoring and alerting systems using OpenTelemetry standards - Participate in on-call rotations, incident response, and post-mortem analyses, ensuring rapid recovery and continuous learning from failures. - Define and track SLOs/SLIs and error budgets to monitor service health and performance. - Implement cloud security best practices to protect sensitive data and maintain the integrity of our systems. - Collaborate across Engineering, Security, and Product teams to embed reliability and automation in every phase of development and deployment. - Contribute to GKE cost optimization and resource management strategies to enhance efficiency and control operational spend. **Requirements** - 4+ years of experience as a SRE/Platform Engineer. - Strong hands-on experience with **GCP** and GKE. - Proficiency in Kubernetes (architecture, deployments, networking, and troubleshooting). - Solid programming or scripting skills in Go, Python, or Bash. - Proficiency with Docker and Linux - Experience with Terraform - Experience with Helm - Experience with GitHub Actions - Strong understanding of monitoring and observability using Prometheus, Grafana, and logging frameworks. - Familiarity with incident management, on-call operations, and post-mortem processes. - Knowledge of network fundamentals (TCP/IP, DNS, Load Balancing). - Experience with PostgreSQL or distributed databases. - Awareness of FinOps and cloud cost management principles. - Excellent problem-solving, communi…