Save application

featherlessai

Jan 22

Machine Learning Engineer — Inference Optimization

Remote Remote · Remote (world)

pytorch researchfulltime

ashby

Match & cover letter

Create a profile to see your match and get a cover letter for this role.

Create profile and match

Job description

You will optimize model inference performance for large-scale machine learning systems. This role involves deep technical work and collaboration with research engineers.

Details

  • Remote work
  • Competitive compensation and meaningful equity
  • Strong experience in ML inference optimization or high-performance ML systems
  • Hands-on experience with PyTorch

The work

  • Optimize inference latency, throughput, and cost for ML models
  • Profile and bottleneck GPU/CPU inference pipelines
  • Collaborate with research engineers to productionize new model architectures
  • Build and maintain inference-serving systems