Save application

snorkelai

3 days ago

Research Scientist - Frontier Benchmarks

Remote Remote · New York City, NY (Hybrid); San Francisco, CA (Hybrid); United States (Remote)

go 316 - research

greenhouse

Match & cover letter

Create a profile to see your match and get a cover letter for this role.

Create profile and match

Job description

You will design next-generation benchmarks and datasets for frontier AI evaluation. This role involves collaboration across teams to translate research insights into actionable data solutions.

Details

  • New York City, NY; San Francisco, CA; or Remote in the United States
  • $200,000—$375,000 USD
  • Ph.D. in machine learning, NLP, or related field preferred; equivalent experience considered
  • Experience in AI/ML evaluation, NLP, or related fields

The work

  • Design datasets for frontier model training and evaluation
  • Translate benchmark insights into narratives for presentations and reports
  • Collaborate with cross-functional teams to inform the company roadmap
  • Stay updated on LLM evaluation research and integrate best practices