Jobedly Post a Job

Technical Product Manager - AI Compute Platform

Nebius · Remote
RemoteFull-timeTechnologySenior$102,000–$138,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Technical Product Manager AI Compute Platform role

Technical Product Manager AI Compute Platform positions focus on delivering results in their domain. This page aggregates open Technical Product Manager AI Compute Platform roles and what employers typically expect.

**About Nebius:** Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. **The role** Our customers build the frontier of AI on top of Nebius — training state-of-the-art models, running production inference at scale, shipping the research and products that define where the field is going next. We are building the AI cloud that the people building the frontier of AI choose deliberately — not on price, not on raw capacity, but on how it works to use it day to day. To do that, we are growing the AI Compute Platform product team and hiring multiple Technical Product Managers across the full surface of the platform. Your scope will be defined by what you bring. We will match your technical strengths, customer experience, and product instincts to the area of the platform where you can have the most impact. The platform is broad — and at our scale, every slice is mission-critical. If you want to help build the best AI cloud in the world — and you have the technical depth to engage engineering leaders as a peer (not as a translator) and the comfort to talk to customers directly — this team is for you. **The platform you'll help build:** - Hardware platforms & launch — bringing new GPU and CPU platforms (GB300, Vera Rubin, ARM/Grace, future generations) to production with full launch readiness across the stack. - Cluster lifecycle & fleet operations — new region launches, 100,000+ GPU cluster bring-up, platform sharding and allocation architecture, release engineering, host-lifecycle automation, operational efficiency. - Reliability & Mission Control — autohealing, health checks, SLA, fault-tolerant training, MTTR reduction, customer trust at scale, observability as a product. - Customer experience & developer surface — Compute APIs, console, CLI, IMDS and in-VM signals, self-service workflows, notifications, customer-facing observability, unified UX across the product line. - GPU & InfiniBand foundational services — drivers, firmware, NCCL, IB/RoCE, NVLink topology, the foundational layer everything else builds on. - Managed runtime platforms — Soperator (Slurm-on-Kubernetes) and MK8S (Managed Kubernetes for AI workloads), powering training and inference for frontier labs. - Platform integrations & emerging workloads — Token Factory integration, RL and agentic workload infrastructure, capacity sharing, new business surfaces as they emerge. - Cross-platform program & delivery — NVIDIA partnership programs, major-maintenance orchestration, cross-stream releases. You will own one of the slices of this platform end-to-end — from strategy and roadmap through delivery, adoption, and measurable outcomes. Your responsibilities will include (regardless of which slice you own): - Own end-to-end product responsibility for your area — strategy, roadmap, discovery, delivery, adoption, measurable customer and platform outcomes. - Design and own the platform contracts customers depend on — APIs, semantics, system events, customer-facing surfaces, operational behavior — at hyperscaler quality. - Drive cross-team execution across platform engineering, networking, storage, Soperator/MK8S, observability, IAM, billing, capacity planning, support, and product design. - Turn customer pain into product commitments through structured discovery — interviews, usage analytics, support patt…

Salary estimate

$102,000 – $138,000/yr
Provided by the employer.

Skills for this role

KubernetesAutomation

Resume tips for Technical Product Manager AI Compute Platform applicants

Interview preparation

Prepare concrete STAR-format stories that show Technical Product Manager AI Compute Platform outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Technical Product Manager AI Compute Platform problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Nebius

Nebius is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles