Cleared Senior Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Cleared Senior Site Reliability Engineer roles and what employers typically expect.
ABOUT GALLATIN At Gallatin, we are rebuilding logistics infrastructure for the national security missions of the United States and allied partners. We build AI systems that determine how logistics decisions are made — not just how they're executed. From factory to foxhole, we operate at the layer where data becomes decisions, and decisions make the advantage. ABOUT THE ROLE Gallatin is looking for a Site Reliability Engineer to keep our production systems running with the reliability our national security customers require. You'll work at the intersection of infrastructure, automation, and mission-critical uptime — building and operating the systems that turn logistics data into decisions in real time, including in classified and disconnected environments. This is a hands-on role for an engineer who treats reliability as a product: someone who instruments before things break, automates the toil away, and owns incidents from detection through postmortem. WHAT YOU'LL DO RELIABILITY & OPERATIONS - Own the reliability, availability, and performance of production systems supporting Gallatin's logistics decision platform, including services deployed in classified and air-gapped environments. - Build and maintain monitoring, alerting, and observability pipelines that surface problems before customers do. - Lead incident response for production issues, driving triage and resolution and running blameless postmortems that turn into concrete engineering fixes. INFRASTRUCTURE & AUTOMATION - Design and operate CI/CD pipelines and infrastructure-as-code that let engineering ship safely and often, across both cloud and classified network environments. - Automate manual operational work (deployments, scaling, failover, credential rotation) to reduce toil and remove single points of failure. - Harden systems to meet DoD security and compliance requirements (e.g., RMF, STIGs, ATO processes) without slowing down delivery. CROSS-FUNCTIONAL PARTNERSHIP - Partner with software engineers to define SLOs/SLIs and build reliability into services from design through deployment. - Work directly with government customers and field teams to understand mission environments and translate operational constraints into system requirements. - Document runbooks, architecture decisions, and operational procedures so the systems you build can be run by the whole team, not just you. WHAT WE’RE LOOKING FOR - Active Secret clearance required; willingness and eligibility to obtain a higher-level clearance if needed. - 3–5 years of experience in a site reliability engineering, DevOps, or production infrastructure role. - Strong technical background in Linux systems, networking, and cloud infrastructure (AWS, Azure, or GovCloud equivalents). - Hands-on experience with infrastructure-as-code (Terraform, Ansible, or similar) and container orchestration (Kubernetes, Docker). - Experience building and maintaining CI/CD pipelines and automated deployment systems. - Familiarity with monitoring and observability tooling (Prometheus, Grafana, Datadog, ELK, or similar). - Track record of owning production incidents end-to-end, from detection through resolution and postmortem. - Comfort operating in classified, air-gapped, or otherwise network-constrained environments. - Clear written and verbal communicator who can work directly with both engineers and government stakeholders. - Willing to travel >50% of the time to customer and government sites. BONUS POINTS - Hands-on experience with Microsoft Azure, including Azure Government (GCC High) or Azure Government Secret / IL5–IL6 environments. - Experience achieving or maintaining an Authority to Operate (ATO) under RMF in a DoD or federal environment. - Experience standing up or operating in multi-cloud environments spanning AWS and Azure. - Familiarity with edge or disconnected/degraded/intermittent/limited (DDIL) deployment patterns for classified or tactical environments. - Prior experience at an early-stage startup or other fas…