Save application

rhoda-ai

May 12

Senior Inference Optimization ML Engineer

Onsite Mountain View

pytorch softwarefulltime

ashby

Match & cover letter

Create a profile to see your match and get a cover letter for this role.

Create profile and match

Job description

You will optimize the performance of large multimodal models for real-world deployment at Rhoda AI. This role involves close collaboration with research and robotics teams to enhance inference efficiency.

Details

  • Mountain View, work from office
  • 3+ years of experience
  • Skills: PyTorch required

The work

  • Own inference performance end-to-end for large foundation models
  • Build systematic performance attribution and identify bottlenecks
  • Apply optimization techniques like quantization and pruning
  • Collaborate with research engineers to implement optimized models