Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)

Perplexity · San Francisco · FullTime

Apply on company site

Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads.

Responsibilities

Qualifications

We expect you to have real depth in most of these:

Additional experience we value

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.