Jobedly Post a Job

Computational Linguist, AI Evaluation

Twelve Labs · San Francisco
Full-timeTechnologyMid Level$140,000–$160,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Computational Linguist AI Evaluation role

Computational Linguist AI Evaluation positions focus on delivering results in their domain. This page aggregates open Computational Linguist AI Evaluation roles and what employers typically expect.

## **Who We Are:** Video is 90% of the world's data. Most of it is invisible to machines. TwelveLabs builds the intelligence layer to change that. Our multimodal AI models understand video the way humans do — across sight, sound, and motion — and power production-scale AI workloads across media, entertainment, sports, security, and government. We have raised more than $210 million from NEA, Radical Ventures, Amazon, NVIDIA, Snowflake, Databricks, Index Ventures, NAVER Ventures, Korea Investment Partners, Quadrille Capital, Red Bull Ventures, and AI pioneers including Fei-Fei Li, Silvio Savarese, and Alexandr Wang. We are a global company, headquartered in San Francisco with offices in Seoul, New York, and London, and employees around the world. We believe the differences in our cultural, educational, and life experiences make our products stronger. Building technology that understands the world in all its complexity requires people who see it from every angle. We are looking for individuals who are driven by hard problems and want their work to matter. Come build it with us! ## About the Role: You will be a vital member of our ML Data Operations Team – which leads the full spectrum of video-language data collection, labeling operations, and model quality measurement. This role comes with high ownership and includes responsibilities such as defining dataset and evaluation requirements in consultation with our research and product teams; designing and building data pipelines; and coordinating with our vendor partners that execute at scale. You will also be responsible for automating as much of the repetitive partnership and annotation-quality-evaluation work as possible. A desire to work cross functionally and to build relationships is critical for success in this position. Position is hybrid in San Francisco - onsite Tuesdays & Thursdays. ## In this role, you will: - **Evaluation Strategy & Design:** Define and build evaluation protocols for video-language model quality, working with Research and Product teams to translate ambiguous quality questions into measurable, repeatable benchmarks or evaluation flows. - **Data-to-Insight Pipelines:** Turn raw customer and usage data into structured signal, building pipelines and analyses that surface where our models are underperforming and working with XFN partners to translate these insights into concrete improvement plans. - **Labeling Operations:** Design and execute video-language data collection and labeling projects, automating repetitive processes so the team can focus on higher-leverage work. - **Vendor & Partner Collaboration:** Coordinate with vendor and outsourcing partners executing at scale, keeping quality high through clear guidelines and feedback loops. - **Cross-Functional Partnership:** Work closely with Engineering and Research teams to align on top-priority data and evaluation needs, communicating findings through dashboards and reports that drive decisions. ## You may be a good fit if you have: - Direct experience building or running model evaluation pipelines (benchmark design, human eval frameworks, automated scoring systems), and translating ambiguous quality questions into measurable criteria. - 5+ years of experience working in an AI focused data operations organization. - A proven track record designing and executing large scale data projects, including gathering, labeling, and post-processing data. - The ability to analyze messy and complex data, identify overarching patterns, and distill your findings into crisp annotation guidelines or other accessible documentation. - Proficiency with Python, agentic coding, or other popular industry tools for automation. - Excellent communication and project management skills, and the ability to support several projects simultaneously. - A foundational understanding of and interest in LLMs/VLMs and multimodal AI. - Conviction that data is the key ingredient for the performance of AI models. ## Preferred Qualifications:…

Salary estimate

$140,000 – $160,000/yr
Provided by the employer.

Skills for this role

PythonSnowflakeProject ManagementCommunicationSecurityAutomation

Resume tips for Computational Linguist AI Evaluation applicants

Interview preparation

Prepare concrete STAR-format stories that show Computational Linguist AI Evaluation outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Computational Linguist AI Evaluation problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Twelve Labs

Twelve Labs is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles