Member of Technical Staff (Software Engineer, Inference & Training Platform)
Perplexity ยท San Francisco ยท Remote (United States) ยท London ยท New York City ยท Ireland ยท posted 4d ago
- Backend
Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers. Today, our inference engineers and researchers build models while also managing networking, securing capacity, and operating the underlying GPU clusters, responsibilities we want a dedicated platform team to own. Your job is to take ownership of that infrastructure and hide its complexity behind a unified, self-serve platform for running training and inference workloads. RESPONSIBILITIES - Build a self-serve compute platform. Design and own the systems that let inference engineers and researchers launch training jobs and operate inference services without managing GPU pro
Listed here from Perplexity careers. Apply opens the employerโs own careers portal. AI CareerPath is not the employer and does not process applications. Never pay a fee to apply for a job.