Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)New$250K–$485K

San FranciscoConstruction

The opportunity

Perplexity serves hundreds of millions of queries a month, and every one of them fans out into multiple AI inference requests running in real time. Behind that sits a large GPU fleet spread across several cloud providers.

What they're looking for

  • Set technical direction across teams. Partner with inference and cloud: infrastructure engineers to turn operational constraints into a coherent platform architecture and roadmap.
  • Deep Kubernetes experience: custom operators, CRDs, and multi-cluster federation, not just running kubectl apply.
  • You've managed GPU clusters at scale: NVIDIA hardware, CUDA, and the networking that makes them fast (InfiniBand or RoCE).
  • You've orchestrated compute across multiple clouds (CoreWeave, AWS, GCP, or: similar) and understand how different each one really is.