Jobedly Post a Job

Site Reliability Engineer

vynca · Remote
RemoteFull-timeEngineeringGeneral$179,000–$241,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Site Reliability Engineer role

Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Site Reliability Engineer roles and what employers typically expect.

Join the dynamic journey at Vynca, where we're passionate about transforming care for individuals with complex needs. We’re more than just a team; we're a close-knit community. Our shared commitment to caring for each other and those we serve is what sets us apart. Guided by our unwavering core values: Excellence, Compassion, Curiosity, and Integrity, we forge paths of success together. Join us in this transformative movement where you can contribute to making a profound difference every day. At Vynca, our mission is to provide comprehensive care for more quality days at home. About the job We're looking for a Site Reliability Engineer (E3) to help build and operate the infrastructure that powers Vynca's healthcare technology platform. In this role, you'll work at the intersection of software engineering, cloud infrastructure, and operations to ensure our systems are reliable, scalable, secure, and performant. As a member of the Technology team, you'll design and manage cloud infrastructure in AWS, operate Kubernetes-based workloads, improve observability across our platform, and automate operational processes that enable engineering teams to move quickly and safely. You'll play a critical role in maintaining the health of our production environment while helping shape the future architecture of our systems. This is a hands-on engineering role with significant ownership and impact. You'll partner closely with Software Engineers, Product teams, and Data teams to build resilient systems that support our mission of delivering comprehensive care for more quality days at home. This position is remote and requires working East Coast business hours (EST). What you'll do - Design, provision, and manage AWS infrastructure using Terraform as the source of truth. - Operate, maintain, and scale production workloads running on Kubernetes. - Package, deploy, and manage applications using Helm and infrastructure automation tools. - Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. - Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity. - Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience. - Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions. - Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability. - Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design. - Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements. - Participate in an on-call rotation and provide operational support for production systems. Your experience and qualifications - Experience: Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, Cloud Infrastructure Engineering, or similar infrastructure-focused roles, preferably within healthcare, SaaS, or high-growth technology environments. - Education: Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related technical field; equivalent professional experience will also be considered. - Strong hands-on experience operating production workloads within AWS environments. - Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation. - Experience operating and supporting production Kubernetes environments. - Hands-on experience deploying and managing applications using Helm. - Experience working with distributed systems, event-driven architectures, or event-sourcing platform…

Salary estimate

$179,000 – $241,000/yr
Provided by the employer.

Skills for this role

AWSKubernetesTerraformDevopsSecurityAutomation

Resume tips for Site Reliability Engineer applicants

Interview preparation

Prepare concrete STAR-format stories that show Site Reliability Engineer outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Site Reliability Engineer problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About vynca

vynca is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles