Protege

Forward Deployed Machine Learning Engineer

Remote, United StatesRemoteFull timePosted 4 days ago
Apply on Protege →

Sign into see who you know at Protege.

Company Overview:

We are building Protege to solve the biggest unmet need in AI — getting access to the right training data. The process today is time intensive, incredibly expensive, and often ends in failure. The Protege platform facilitates the secure, efficient, and privacy-centric exchange of AI training data.

Solving AI’s data problem is a generational opportunity. We’re backed by world-class investors and already powering partnerships with some of the most ambitious teams in AI. The company that succeeds will be one of the largest in AI — and in tech.

We’re a lean, fast-moving, high-trust team of builders who are obsessed with velocity and impact. Our culture is built for people who thrive on ambiguity, own outcomes, and want to shape the future of data and AI.

About the Role

We're hiring a Forward Deployed Machine Learning Engineer in our Benchmarks and Evaluations vertical. You'll be the first MLE dedicated to this vertical and will work directly with the GM and our researchers to scale Protege’s position as a renowned leader in the space.

At Protege, we believe that real world data is one of the largest bottlenecks to AI progress. Our data and data expertise position us to be neutral arbiters for the market, helping model builders understand the current performance of their models, identify what data will improve performance, and show that improvement over time. Benchmarks and evaluations power that cycle. As an early engineer in the Benchmarks and Evaluations vertical, this role is an opportunity to help build the technical foundation for a critical area that greatly benefits current and future customers.

What You'll Do

Work on the eval foundation

• Partner with the GM and early customers to define what constitutes strong evals in different domains

• Work with Protege researchers to design and build benchmarks

• Build the standards on how different modalities should be processed

Own infrastructure

• Build the backend the vertical...