Jobedly Post a Job

Staff SRE - Observability

Focused · Chicago, IL
Full-timeEngineeringTechnology$179,000–$241,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Staff Sre Observability role

Staff Sre Observability positions focus on delivering results in their domain. This page aggregates open Staff Sre Observability roles and what employers typically expect.

Who we are: At Focused, we move quickly to deliver quality software that achieves client outcomes and meets their customer’s needs. We strategically partner with our clients to leverage our expertise in design and software, while our clients bring their own domain expertise. We work with a variety of clients from different industries, collaborating as we get new products to market, modernizing legacy systems, or helping teams learn the skills they need to be successful. Our values: Listen first • We are experts in product practices but life long learners in the domain of our customers. We research, collaborate, and understand. Learn why • We ask questions and talk to users to understand problem spaces, objectives, and goals, which allows us to deeply invest and drive towards the outcomes of our clients. Love your craft • We love diving into a variety of domains and solving problems. We take pride in delivering value, in communicating progress, and guiding our clients to success. We are seeking an experienced Staff Observability Consultant with deep expertise in OpenTelemetry and strong Platform Engineering capabilities to help organizations implement, optimize, and scale their observability infrastructure. This role requires a seasoned consultant who can design comprehensive telemetry strategies, implement distributed tracing solutions, establish robust monitoring practices, and interface closely with clients on the observability journey. Key Responsibilities: OpenTelemetry & Observability Design and implement end-to-end OpenTelemetry solutions across diverse technology stacks Configure and deploy OpenTelemetry Collectors for efficient data collection, processing, sampling, and routing Establish telemetry pipelines for metrics, traces, and logs across microservices architectures Optimize collector configurations for performance, reliability, and cost-effectiveness Platform Engineering & Infrastructure Augment existing infrastructure with with integrated observability solutions Implement Infrastructure as Code (IaC) solutions using Terraform, Pulumi, CloudFormation, etc. Architect and manage Kubernetes clusters with comprehensive monitoring and logging Build CI/CD pipelines with embedded observability and automated testing Site Reliability Engineering (SRE) Establish and maintain Service Level Indicators (SLIs), Objectives (SLOs), and Agreements (SLAs) Implement error budgets, toil reduction strategies, and capacity planning Support incident response procedures and post-mortem processes Cloud & DevOps Engineering Deploy and manage observability infrastructure across AWS, GCP, and Azure Establish security, compliance, and governance frameworks for telemetry data Experience automating Agent Evaluations in CI/CD pipelines and observability backends. Required Qualifications: Core Observability & OpenTelemetry 3-7 years of experience in observability, monitoring, and distributed systems Deep hands-on experience with OpenTelemetry ecosystem, including SDKs, APIs, and specifications Proficiency with OpenTelemetry Collector configuration, processors, exporters, and receivers Strong understanding of telemetry data models, semantic conventions, and instrumentation best practices Platform Engineering & DevOps 5+ years of Platform Engineering or DevOps experience with focus on site reliability, observability, and incident response Proficiency with Infrastructure as Code tools (Terraform, Pulumi, CloudFormation, CDK) Strong experience with CI/CD platforms (GitHub Actions, GitLab CI, Jenkins, ArgoCD) Cloud & Infrastructure Hands-on experience with major cloud providers (AWS, GCP, Azure) and their observability services Experience with container technologies (Docker, Podman) and container registries Knowledge of networking, security, load balancing, and distributed systems concepts Site Reliability Engineering Experience implementing SRE practices including error budgets and toil metrics Proficiency in incident management, on-call procedures…

Salary estimate

$179,000 – $241,000/yr
Provided by the employer.

Skills for this role

AWSAzureGCPDockerKubernetesTerraformCi/CdDevopsSecurity

Resume tips for Staff Sre Observability applicants

Interview preparation

Prepare concrete STAR-format stories that show Staff Sre Observability outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Staff Sre Observability problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Focused

focused is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles