Save application

hyperbolic

26 days ago

Member of Technical Staff - Inference

Onsite San Francisco, CA

kubernetes engineeringfulltime

ashby

Match & cover letter

Create a profile to see your match and get a cover letter for this role.

Create profile and match

Job description

You will build inference capabilities for our AI platform in San Francisco. This role involves deploying and optimizing models across global clusters using Kubernetes.

Details

  • San Francisco, work from office
  • Experience with Kubernetes required

The work

  • Deploy models on Forge and Kubernetes
  • Evaluate inference frameworks and set up monitoring
  • Optimize and debug customer inference
  • Build end-to-end product features