Jobedly Post a Job

Senior Site Reliability Engineer (SRE)

UJET · Seoul
Full-time000001 - EngineeringTechnology$179,000–$241,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Senior Site Reliability Engineer role

Senior Site Reliability Engineer positions focus on delivering results in their domain. This page aggregates open Senior Site Reliability Engineer roles and what employers typically expect.

About Us UJET leads the way in AI-powered contact center innovation, delivering a future-proof, cloud platform that redefines the customer experience with cutting-edge AI, true multimodality, and a mobile-first approach. We infuse AI across every aspect of your customer journey and contact center operations, to drive automation and efficiency. UJET's AI solutions empower agents, optimize customer journeys, and transform contact center operations for elevated experiences and actionable insights. Built on a cloud-native architecture with a unique CRM-first approach, UJET ensures unmatched security, scalability, and prioritized data insights (without storing PII). Designed for effortless use, UJET partners with businesses to deliver exceptional interactions, smarter decision-making, and accelerated growth in the AI-driven world. Learn more at www.ujet.cx . Opportunity We’re looking for a Senior Site Reliability Engineer to help build and scale a high-impact SRE function. You’ll be a technical leader on a team responsible for improving system reliability, reducing operational toil, and establishing best practices across engineering. In this position, you’ll design how reliability works in UJET, influence engineering decisions, and build the tooling and processes that make production safer and more predictable. Responsibilities Lead efforts to improve system reliability, scalability, and performance across critical services Define and implement SLIs/SLOs and error budgets, and use them to guide engineering priorities Design and develop observability systems (metrics, logging, tracing, alerting) that produce actionable alerts and data. Lead complex incident response, acting as incident commander when needed Conduct postmortems focused on systemic causes rather than individual fault, and ensure corrective actions from those reviews are completed. Identify and eliminate toil through automation, tooling, and improved workflows Partner with product and platform teams on architecture decisions, production readiness, and designing systems that recover from failure Build reusable systems and “paved roads” that make it easier for teams to operate their services reliably Mentor other engineers and raise the overall operational maturity of the organization Requirements Must be a South Korean citizen or currently reside in South Korea with a valid work visa 6-10+ years of experience in SRE, infrastructure, or backend systems engineering Demonstrated experience of owning reliability outcomes for complex, distributed systems Strong experience with cloud infrastructure (AWS, GCP, or Azure) and production-scale systems Deep understanding of observability, incident management, and system performance Proficiency in at least one programming language (e.g., Go, Python, Java) with a focus on automation and tooling Able to change how other teams work without having managerial authority over them Strong competency in making clear decisions during incidents by following a defined process without reacting emotionally. Stand Out Qualifications Experience building or scaling SRE practices (SLOs, incident frameworks, on-call models) Kubernetes/container orchestration experience Infrastructure as Code (Terraform, etc.) Experience with high-growth or scaling systems Background in performance engineering or capacity planning Success Criteria Critical services have clear, meaningful SLOs that drive engineering decisions Alerts are actionable; irrelevant alerts are reduced; on-call workload is manageable. Incidents are handled efficiently, and repeat issues decline over time Engineering teams adopt reliability best practices with minimal friction Toil is actively reduced through automation and better system design Position Context This is an early position in the company's SRE function. You will have direct input into how reliability standards and practices are established, which forms the foundation on which product engineering builds. UJET is changing how compa…

Salary estimate

$179,000 – $241,000/yr
Provided by the employer.

Skills for this role

PythonJAVAGOAWSAzureGCPKubernetesTerraformSecurityAutomation

Resume tips for Senior Site Reliability Engineer applicants

Interview preparation

Prepare concrete STAR-format stories that show Senior Site Reliability Engineer outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Senior Site Reliability Engineer problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About UJET

UJET is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles