Jobedly Post a Job

Senior Principal Software Engineer, Machine Learning

Toast · Remote
RemoteFull-timeR & D : Foundations : AITechnology$179,000–$241,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Senior Principal Software Engineer Machine Learning role

Senior Principal Software Engineer Machine Learning positions focus on delivering results in their domain. This page aggregates open Senior Principal Software Engineer Machine Learning roles and what employers typically expect.

Toast creates technology to help restaurants and local businesses succeed in a digital world, helping business owners operate, increase sales, engage customers, and keep employees happy. About the Team & Role: The Machine Learning Platform team builds and operates the core infrastructure that powers AI and ML across Toast — the feature store, model hosting and serving, the experimentation platform, training pipelines, and the tooling ML engineers and data scientists rely on every day. Our work directly enables the models that drive personalization, forecasting, fraud detection, search, and the growing set of AI-powered experiences shipping to restaurants. Toast is seeking a Senior Principal Software Engineer to act as the technical leader on the ML Platform team, shaping the systems that will carry Toast's AI and ML capabilities into the next decade. The role involves driving architectural direction across the platform, delivering foundational infrastructure that other teams build on, and elevating fellow engineers. The ideal candidate is a domain expert with considerable ML platform experience, who partners with ML engineers, data scientists, product, and infrastructure leadership on high-leverage opportunities. This position suits an engineer comfortable writing production code, leading technical design for distributed systems, and influencing organizational decisions about how Toast builds and deploys ML. A day in the life (Responsibilities): Own technical direction of the ML Platform — feature store, model hosting and serving, experimentation, training infrastructure — driving architectural decisions around scalability, reliability, latency, and cost Lead design and delivery of large-scope platform initiatives from conception through production, coordinating across ML, data, and infrastructure teams Identify and resolve systemic technical challenges: online/offline feature parity, model deployment friction, experimentation velocity, GPU utilization, cross-team dependencies Set and maintain a high engineering quality bar through hands-on code contributions, design reviews, and mentorship of platform and ML-adjacent engineers Partner with ML engineering, data science, product, and platform leadership to translate ML strategy into technical roadmaps Define the paved paths ML teams use to ship models safely — from feature registration through canary rollout, monitoring, and rollback Leverage AI-augmented development tools to increase development velocity and code quality What you'll need to thrive (Requirements): 10+ years delivering complex backend or infrastructure systems at scale Direct experience building or operating core ML infrastructure — feature stores, model serving, experimentation platforms, training orchestration, or equivalent Mastery of a modern backend language, ideally Java or Kotlin Deep proficiency with distributed systems concepts: consistency, latency, throughput, fault tolerance, and observability Strong understanding of data modeling, query languages, and the online/offline data patterns that underpin ML systems Demonstrated technical leadership, with ability to drive cross-team alignment and influence engineering, product, and business stakeholders Bachelor's degree in Computer Science or a related field, or equivalent practical experience What will help you Stand Out (Nice to Haves): Hands-on experience with open-source or commercial ML platform components (e.g. Tecton, MLflow, SageMaker, Databricks) Experience building or operating experimentation / A-B testing platforms at scale Familiarity with real-time streaming systems (Kafka, Flink, Spark Streaming) and their use in feature computation Experience serving LLMs or deep-learning models in production, including GPU capacity planning and inference optimization Prior work supporting internal-developer-facing platforms with a product mindset AI at Toast At Toast, one of our company values is that we're hungry to build and learn. We believe learning n…

Salary estimate

$179,000 – $241,000/yr
Provided by the employer.

Skills for this role

JAVASparkKafkaMachine LearningSalesLeadershipData ScienceKotlin

Resume tips for Senior Principal Software Engineer Machine Learning applicants

Interview preparation

Prepare concrete STAR-format stories that show Senior Principal Software Engineer Machine Learning outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Senior Principal Software Engineer Machine Learning problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Toast

Toast is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles