Twelve Labs

Computational Linguist, AI Evaluation

San Francisco, California, United StatesFull time$140,000 - $160,000 / yearPosted 13 days ago
Apply on Twelve Labs →

Sign into see who you know at Twelve Labs.

WHO WE ARE:

Video is 90% of the world's data. Most of it is invisible to machines.
TwelveLabs builds the intelligence layer to change that. Our multimodal AI models understand video the way humans do — across sight, sound, and motion — and power production-scale AI workloads across media, entertainment, sports, security, and government.

We have raised more than $210 million from NEA, Radical Ventures, Amazon, NVIDIA, Snowflake, Databricks, Index Ventures, NAVER Ventures, Korea Investment Partners, Quadrille Capital, Red Bull Ventures, and AI pioneers including Fei-Fei Li, Silvio Savarese, and Alexandr Wang.

We are a global company, headquartered in San Francisco with offices in Seoul, New York, and London, and employees around the world. We believe the differences in our cultural, educational, and life experiences make our products stronger. Building technology that understands the world in all its complexity requires people who see it from every angle. We are looking for individuals who are driven by hard problems and want their work to matter. Come build it with us!

ABOUT THE ROLE:

You will be a vital member of our ML Data Operations Team – which leads the full spectrum of video-language data collection, labeling operations, and model quality measurement. This role comes with high ownership and includes responsibilities such as defining dataset and evaluation requirements in consultation with our research and product teams; designing and building data pipelines; and coordinating with our vendor partners that execute at scale. You will also be responsible for automating as much of the repetitive partnership and annotation-quality-evaluation work as possible. A desire to work cross functionally and to build relationships is critical for success in this position.

Position is hybrid in San Francisco - onsite Tuesdays & Thursdays.

IN THIS ROLE, YOU WILL:

- Evaluation Strategy & Design: Define and build evaluation protocols for video-language model qua...