Jobedly Post a Job

Senior Observability Engineer

FanDuel · New York City
Full-timePlatform EngineeringTechnology$136,000–$184,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Senior Observability Engineer role

Senior Observability Engineer positions focus on delivering results in their domain. This page aggregates open Senior Observability Engineer roles and what employers typically expect.

THE POSITION Our roster has an opening with your name on it FanDuel is looking for a Senior Observability Engineer to design, build, and mature the observability ecosystem that underpins our platform and services. You will deliver deep visibility into system behavior by combining system telemetry with user signals to provide a holistic view of performance, reliability, and user experience. You’ll also explore how AI and machine learning can enhance observability, from intelligent alerting and anomaly detection to accelerating root cause analysis. This is a hands-on role. You’ll partner closely with engineering and product teams to deliver scalable observability capabilities, serve as a subject matter expert in monitoring, alerting, and incident management, and equip teams with self-service insights and tooling. By connecting system behavior to real user impact and leveraging AI-assisted workflows to surface issues faster, you’ll drive improvements in reliability, performance, and data-informed decision-making across the organization. In addition to the specific responsibilities outlined above, employees may be required to perform other such duties as assigned by the Company. This ensures operational flexibility and allows the Company to meet evolving business needs. THE GAME PLAN Everyone on our team has a part to play Contribute to the observability strategy and roadmap, partnering with multiple teams to align with business priorities and engineering goals. Design and enhance scalable observability solutions that provide actionable insights into system health, performance, and user experience. Help establish and promote best practices for monitoring, alerting, incident management, and postmortems across teams. Support operational excellence by improving incident response processes, on-call practices, and post-incident reviews, focusing on continuous improvement. Collaborate on cross-team initiatives to improve system reliability, identifying risks and contributing to their resolution. Apply automation and AI-assisted workflows to improve root cause analysis and reduce operational toil. Work with engineering and product stakeholders to surface observability insights that inform technical decisions and prioritization. Analyze system and user signals to help detect, prevent, and mitigate reliability issues. Contribute to optimizing observability platforms for performance, scalability, and cost-efficiency. Mentor peers and contribute to raising observability and reliability standards within the team. In addition to the responsibilities outlined above, employees may be required to perform other duties as assigned by the Company to ensure operational flexibility and meet evolving business needs. A Sneak Peek Into Our Tech Stack AWS, Kubernetes, Terraform, Helm, Ansible, Vault, Datadog and PagerDuty THE STATS What we're looking for in our next teammate Solid hands-on experience in observability engineering, SRE, platform engineering, or related roles, with impact across team-level systems. Strong expertise in monitoring and observability practices, with hands-on experience using tools such as Datadog. Experience contributing to observability or reliability initiatives across teams or services. Proficiency with Kubernetes, cloud infrastructure (e.g. AWS), and infrastructure-as-code tools such as Terraform. Ability to influence technical decisions within and across teams, collaborating effectively with a range of stakeholders. Good understanding of distributed systems principles (e.g. consistency, availability, partition tolerance) and practical trade-offs. Experience defining and implementing SLOs, SLIs, and alerting strategies, including an understanding of user-impacting metrics. Strong software engineering fundamentals, with proficiency in at least one modern programming language (e.g. Go, Java, Python, or TypeScript), and experience building tooling, automation, and scalable systems. Experience improving systems through automati…

Salary estimate

$136,000 – $184,000/yr
Provided by the employer.

Skills for this role

TypescriptPythonJAVAGOAWSKubernetesTerraformMachine LearningAutomation

Resume tips for Senior Observability Engineer applicants

Interview preparation

Prepare concrete STAR-format stories that show Senior Observability Engineer outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Senior Observability Engineer problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About FanDuel

FanDuel is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles