Member of Technical Staff (Software Engineer, Inference & Training Platform)

Perplexity · San Francisco · FullTime

Apply on company site

Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads.

Responsibilities

Qualifications

We expect you to have real depth in most of these:

Additional experience we value

If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above.