APPLY TO SPEEDRUN
← Preference Model
Preference Model · Hiring

Member of Technical Staff - Research & Post-training

San FranciscoOn SiteFull Time$200K – $350K • Offers Equity

About Us

Preference Model is building automated ML research engineering.

Existing frontier models are brittle when applied to real-world ML tasks. The present bottleneck is the lack of high-quality RL training environments. Our first step is to build RL environments that reflect real-world complexity, with diverse tasks and robust reward functions.

Our founding team has previous experience on Anthropic’s data team building data infrastructure, and datasets behind Claude. We are partnering with leading AI labs to push AI closer to achieving its transformative potential.

About the Role

Models of the future will be able to train themselves on tasks that they are not good at. We are interested in investigating how far we can push the boundaries of self-directed learning. We are looking for Research Engineers or Research Scientists to push the frontier of post-training on large language models in a role that blends research and engineering, requiring you to implement novel approaches and shape research directions.

What You Will Do:

What We are Looking For

You may be a good fit if you also:

Candidates don't need a PhD or extensive publications. Some of the best researchers have no formal ML training and gained experience building industry products. We believe adaptability combined with exceptional communication and collaboration skills are the most important ingredients for successful startup research.

What We Offer:

We value diverse perspectives and experiences. If you're excited about this role but don't check every box, we still encourage you to apply.

Interested in This Role?

Apply at Preference Model

You'll head to Preference Model's own careers page.