Save application

Artificial Analysis

4 days ago

Member of Technical Staff (MTS), Language Model Evaluations

Onsite San Francisco, CA USA

hacker newswho is hiring

hnwhoishiring

Match & cover letter

Create a profile to see your match and get a cover letter for this role.

Create profile and match

Job description

You will work on building evaluation datasets and benchmarks for language models in San Francisco. This role involves hands-on technical tasks to enhance model performance.

Details

  • San Francisco, work from office

The work

  • Build frontier evals including datasets and benchmarks
  • Develop harnesses for evaluations