All open roles

Member of the Technical Staff

Engineering · Posted 2 months ago

We are looking for a Member of the Technical Staff with 5+ years of combined academic and industry research experience to lead the research behind evaluations and datasets that will enable the next wave of AI model improvement. You'll be one of the first dedicated research hires, shaping how the company thinks about da...

Way of working
Hybrid
Location
New York, NY, San Francisco, CA
Pay range
$250,000 to $350,000
Level
Staff
Experience
6 to 10 years
Type
Full time
Visa sponsorship
Not offered for this role
The company
20 to 50 people

Skills that matter here

PythonPyTorchLLMsEvaluation FrameworksData Pipelines

The full description

We are looking for a Member of the Technical Staff with 5+ years of combined academic and industry research experience to lead the research behind evaluations and datasets that will enable the next wave of AI model improvement. You'll be one of the first dedicated research hires, shaping how the company thinks about data quality, LLM evaluation, and benchmark design across critical domains like life sciences and finance.

What will you be doing?

- Designing and owning benchmarks, evals, and dataset taxonomies that set the standard for data quality in AI

- Conducting independent research and publishing papers to establish the company as a thought leader in AI benchmarks

- Identifying where frontier models are failing and building datasets to address those gaps

- Partnering with domain experts and deployment leads to define what high-quality training and evaluation data looks like across verticals

- Driving the research agenda with autonomy — proposing new ideas, designing experiments, and dragging the team along with you

Key Requirements

- PhD in machine learning, computer science, computer vision, or a quantitative science (strongly preferred)

- Published research at top venues (e.g., ICML, NeurIPS) — ideally as a top author with cited work

- Deep familiarity with LLMs, benchmark design, eval methodology, and data quality for model improvement

- Based in New York or San Francisco and willing to work in-office 5 days/week (flexible to 4 days)

- Natural curiosity about model behavior with the ability to identify where AI is deficient and propose data-driven solutions

Interested in this one?

There is no apply button here on purpose. Tell us about yourself, we book a short call, and if this role fits we walk you through the company and ask before anything is sent. Always free for you.

Tell us about you
Tell us about you