Jobedly Post a Job

Machine Learning Architect - Conversational Speech

Apple · Cupertino, CA
Full-timeGeneralSenior$187,000–$253,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Machine Learning Architect Conversational Speech role

Machine Learning Architect Conversational Speech positions focus on delivering results in their domain. This page aggregates open Machine Learning Architect Conversational Speech roles and what employers typically expect.

Join the team redefining what a deeply personal and integrated assistant can be. As part of the Siri organization, you will help shape one of the world's most widely used AI assistants, powered by our next-generation of Apple Intelligence, with capabilities like personal context understanding and on-screen awareness, built with privacy from the ground up. Your work will have direct, meaningful impact for users across iOS, iPadOS, macOS, watchOS, and visionOS. This is a rare opportunity to build at the intersection of cutting-edge AI and human-centered design, shipping technology that is centered around users and their needs. ## Description The Speech organization within Siri drives major speech recognition, synthesis, and speech-to-speech model advances for features deeply embedded throughout Apple's ecosystem. Our mission is to build cutting-edge infrastructure, datasets, and models that empower Siri conversational AI, dictation, and speech-enabled Apple Intelligence features across natural language understanding, dialog generation, speech recognition, and multimodal interaction. We apply these technologies to create engaging, intelligent, and personalized conversational experiences for millions of Apple users. We are seeking a Machine Learning Architect to serve as a senior technical leader spanning the full Speech organization. You will set the future modeling direction for all of conversational speech—charting the architectural and algorithmic course for how Apple's speech technologies evolve. You will operate as a hands-on expert who not only defines strategy but also digs into the hardest technical problems, working shoulder-to-shoulder with teams to overcome critical obstacles. Reporting directly to the Speech organization leadership, you will have broad visibility and influence across speech recognition, synthesis, dialog, multimodal foundation models, and speech-to-speech systems, ensuring coherent technical vision and cross-team alignment. ## Responsibilities As the Machine Learning Architect for Conversational Speech, you will define modeling strategy and technical direction across the Speech organization, establishing a unified architectural vision for speech recognition, speech synthesis, dialog systems, multimodal foundation models, and speech-to-speech technologies. You will serve as the organization's foremost modeling expert, providing deep technical guidance to multiple teams working on interconnected speech capabilities. You will evaluate emerging research and industry trends—including advances in large language models, multimodal architectures, and full-duplex natural conversational systems—and translate them into actionable roadmaps. You will champion production-readiness, ensuring architectural decisions account for on-device constraints, latency, scalability, and robustness. You will collaborate broadly with partner teams across Siri, Apple Intelligence, hardware, and platform engineering to ensure speech modeling investments are well-integrated into Apple's broader AI strategy. ## Minimum qualifications 10+ years of experience in machine learning applied to speech or multimodal systems, with progressively increasing technical scope and leadership. Demonstrated expertise as a technical leader or architect who has defined modeling direction across multiple teams or product areas. Deep, hands-on proficiency in modern deep learning, including large language models and end-to-end speech systems. Significant experience with multimodal LLMs, including architecture design, training, adaptation, and deployment of models that integrate speech, audio, and text modalities. Direct experience building speech-to-speech conversational systems, with a strong understanding of full-duplex natural conversational interaction and end-to-end speech pipelines. A track record of translating research into production-quality systems at scale. Expert programming skills in Python and deep learning frameworks such as PyTorch, JAX,…

Salary estimate

$187,000 – $253,000/yr
Provided by the employer.

Skills for this role

PythonPytorchMachine LearningLeadership

Resume tips for Machine Learning Architect Conversational Speech applicants

Interview preparation

Prepare concrete STAR-format stories that show Machine Learning Architect Conversational Speech outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Machine Learning Architect Conversational Speech problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Apple

Apple is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles