Jobedly Post a Job

Staff Site Reliability Engineer

Diligent Corporation · New York, NY
Full-timeEngineering GroupHospitality & Food$179,000–$241,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Staff Site Reliability Engineer role

Staff Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Staff Site Reliability Engineer roles and what employers typically expect.

Role Overview You’re a seasoned Site Reliability Engineer who loves owning complex infrastructure, making things run faster, safer, and with less manual effort. In this Staff‑level role, you’ll design and operate VMware‑based private cloud platforms that power mission‑critical SaaS products used by customers around the world. You’ll work across Linux, Windows Server, networking, storage, and automation frameworks to increase reliability, reduce toil, and modernize a global datacenter environment. You’ll have the scope to set technical direction, build automation at scale, and mentor engineers while staying hands‑on with VMware vSphere, F5/AVI load balancers, and hybrid Active Directory. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead the architecture, deployment, and ongoing optimization of VMware vSphere–based private cloud infrastructure across multiple global datacenters. Design and build automation using PowerShell/PowerCLI, Ansible, Python, and CI/CD tools to streamline provisioning, configuration, and compliance. Administer, harden, and troubleshoot Linux (RHEL/CentOS/Ubuntu) and Windows Server environments that host enterprise and SaaS workloads. Integrate and manage Active Directory for authentication, access control, and service accounts across hybrid on‑prem and cloud environments. Partner with network and security teams to manage firewalls, VPNs, storage, and load balancers (F5 BIG‑IP, AVI/NSX Advanced Load Balancer) for highly available services. Document architectures and runbooks, participate in on‑call and change management, and mentor engineers while influencing long‑term reliability and automation strategy. These are the essentials you’ll need to get an interview 10+ years of experience in systems or infrastructure engineering, including operating large‑scale enterprise or SaaS datacenter environments. Deep hands‑on expertise with VMware vSphere (ESXi, vCenter, DRS, HA, vMotion, distributed switches) in production. Strong Linux administration skills (RHEL/CentOS/Ubuntu), including performance tuning, system hardening, and advanced troubleshooting. Solid experience with Windows Server and Active Directory (Group Policy, DNS, authentication and access integrations). Proven track record building and maintaining automation using PowerShell/PowerCLI, Ansible, Python, or similar tools, plus familiarity with Git or other version control. Good understanding of storage (SAN/NAS), TCP/IP networking, DNS, VPNs, firewalls, and production monitoring/alerting. A collaborative, problem‑solving mindset with the ability to lead complex incidents, communicate clearly, and operate in an on‑call, high‑availability environment. It would be great if you had these too, but we’ll support you if you don’t Experience with enterprise storage and compute platforms such as Pure Storage or Cisco UCS. Familiarity with Terraform, Jenkins, Azure DevOps, or similar tools for infrastructure as code and CI/CD automation. Exposure to security hardening and compliance frameworks such as CIS benchmarks, NIST, or ISO 27001. U.S pay range $131,000 — $164,000 USD About Us Diligent is the AI leader in governance, risk and compliance (GRC) SaaS solutions, helping more than 1 million users and 700,000 board members to clarify risk and elevate governance. The Diligent One Platform gives practitioners, the C-Suite and the board a consolidated view of their entire GRC practice so they can more effectively manage risk, build greater resilience and make better decisions, faster. At Diligent, we're building the future with people who think boldly and move fast. Whether you're designing systems that leverage large language models or part of a team reimaging workflows with AI, you'll help us unlock entirely new ways of working and thinking. Curiosity is in our DNA, we look for individuals willing to ask the big questions and experiment fearlessly - those who embrace change not as a challenge, but as an opportunity. The future…

Salary estimate

$179,000 – $241,000/yr
Provided by the employer.

Skills for this role

PythonAzureTerraformCi/CdGITDevopsSecurityAutomation

Resume tips for Staff Site Reliability Engineer applicants

Interview preparation

Prepare concrete STAR-format stories that show Staff Site Reliability Engineer outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Staff Site Reliability Engineer problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Diligent Corporation

Diligent Corporation is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles