Senior Engineer positions focus on delivering results in their domain. This page aggregates open Senior Engineer roles and what employers typically expect.
This is a U.S. based position. All of the programs we support require U.S. citizenship to be eligible for employment. All work must be conducted within the continental U.S. Who we are: Raft ( https://TeamRaft.com ) is a customer-obsessed non-traditional defense tech company dedicated to empowering U.S. military and government agencies with cutting-edge AI/ML and data solutions. We are a leader in autonomous data fusion and Agentic AI, with a purposeful focus on Distributed Data Systems, Platforms at Scale, and Complex Application Development. With headquarters in McLean, VA, our range of clients includes innovative federal and public agencies leveraging design thinking, cutting-edge tech stack, and cloud-native ecosystem. We build digital solutions that impact the lives of millions of Americans. About the role: As a Senior Engineer, you will design, implement, and maintain secure, scalable platform infrastructure supporting mission-critical applications across cloud, on-premises, and hybrid environments. You will lead the automation of software delivery, infrastructure provisioning, and security integration while partnering closely with software engineers, cybersecurity teams, and customer stakeholders. This role requires deep expertise in Kubernetes, cloud-native technologies, Infrastructure as Code, CI/CD automation, and GitOps practices. You will be responsible for building resilient platforms, improving developer productivity, and ensuring secure, compliant application delivery throughout the software lifecycle. What You Will Do: Design, develop, and maintain Infrastructure as Code ( IaC ) supporting distributed, containerized, and cloud-native application lifecycles. Architect, deploy, and manage Kubernetes platforms across AWS and other cloud environments. Design, implement, and optimize secure, performant CI/CD pipelines that automate testing, security scanning, deployment, and release management. I mplement and mature GitOps workflows using modern GitOps tooling for continuous software delivery. Deploy, configure, and support platform hosted services in support of DOW tenant application teams across multiple impact levels. Build and maintain highly available , production-grade Kubernetes clusters while ensuring compliance with security and operational standards. Securely provide multi-tenant access to cloud native services. Develop automation to improve platform reliability, scalability, and operational efficiency. Configure and maintain Kubernetes networking technologies including Calico, Cilium, or equivalent Container Network Interfaces (CNIs). Deploy and support observability platforms using Prometheus, Grafana, Fluent Bit, Loki, Kibana, and related monitoring and logging solutions. Collaborate with software, platform, and security engineers to integrate security throughout the development lifecycle. What we are looking for: 5+ years of professional experience in DevOps, DevSecOps, Platform Engineering, Site Reliability Engineering, or a related field. 5+ years of hands-on experience deploying, operating, and securing production Kubernetes and Docker environments. Experience provisioning, hardening, and maintaining compliant Kubernetes clusters. Strong understanding of Kubernetes architecture, Helm Charts, and container networking technologies including Calico, Cilium, or similar solutions. Experience operating stateful, highly available, and distributed services in Kubernetes environments. Experience with AWS or other cloud platforms supporting enterprise workloads. Experience with Infrastructure as Code using Terraform, Ansible, or similar automation frameworks. Experience configuring and maintaining CI/CD pipelines using GitLab Runners. Experience with software packaging, artifact repositories, and build automation for distributed applications. Experience implementing monitoring and logging solutions using Prometheus, Grafana, Fluent Bit, Loki, Kibana, or similar observability tools. Strong troubleshooting and…