Senior Devops Engineer positions focus on delivering results in their domain. This page aggregates open Senior Devops Engineer roles and what employers typically expect.
We're hiring someone who wants production infrastructure that has to hold at 3am, every time, for millions of traders who never stop trading. Real uptime targets. Real incidents. Real consequences when something breaks. You'll own that infrastructure end to end, not just the parts that are already stable. * * * **Why This Matters** Trading for Anyone, Anywhere, Anytime means services that don't sleep, across time zones and regulatory regimes. That scale doesn't run on tribal knowledge and manual runbooks. We're already running AI-assisted monitoring that catches problems before they page anyone, automation that remediates known failure patterns on its own, and security workflows that flag threats faster than a human scanning dashboards ever could. Not experiments. Systems carrying production traffic right now. You won't be maintaining this from a distance. You'll be designing it, breaking it, fixing it, and deciding what gets built next. * * * **Why Deriv** We're in production, not planning. - Autonomous security analysts already triaging alerts and correlating threats against historical patterns - Dozens of fraud detection models running continuously against real transactions - Automated security review on every pull request, every day - Infrastructure-as-code, monitoring logic, and runbooks increasingly generated, tested, and documented with AI as a normal part of the workflow, not a side experiment We share what we learn. [Deriv](https://derivai.substack.com) is where we write about AI in production, including what breaks and what we figured out the hard way. You'll own systems that are already handling real transactions, not prototypes waiting for product-market fit. * * * **What You’ll Do** This role owns outcomes across production infrastructure and reliability engineering, with regular cross-functional work alongside the security team: - **Production Infrastructure** — Cloud, container, database, monitoring, and CI/CD environments for high-availability services. You design it, not just patch it. - **AI-Native Delivery** — Automation design, infrastructure-as-code generation, testing, refactoring, documentation, and runbook creation, with AI tooling built into how the work gets done. - **Incident Response & Resilience** — Alerting logic, remediation scripts, self-healing patterns, circuit breakers, and fault-tolerant architecture for systems that can't afford to go down. - **Security Operations** — Hardening, intrusion detection, configuration audits, and AI-enhanced threat detection, built jointly with the security team. - **End-to-End Ownership** — Take a vague operational problem, design the architecture, implement the solution, deploy it safely, monitor the outcome, and keep iterating once it's live. - **Monitoring & Observability** — Maintain and extend Datadog, Grafana, and custom observability systems, with AI-assisted analysis to catch anomalies and potential failures early. - **Incident Diagnosis** — Diagnose and resolve production incidents across complex systems, coordinate with developers, and turn what you learn into durable infrastructure improvements, not one-off fixes. - **Autonomous Operation**s — Explore, prototype, and deploy autonomous AI systems for operations, including always-on agents that maintain context, investigate anomalies, and take approved remediation actions. * * * **Who You Are** You work across all three paradigms that make production infrastructure actually reliable: - **Deterministic systems** — the Terraform, the CI/CD pipeline, the database that has to be right every time - **Predictive systems** — the anomaly detection that flags "this looks wrong" before a human would notice - **Agentic systems** — AI tools that draft the automation, the docs, the first pass at a fix, so your judgment goes toward the decisions that actually need it **Required Experience** - 4+ years of DevOps, SRE, infrastructure engineering, or production operations experience - Hands-on experience shipping and…