Member of Technical Staff (Software Engineer, Cloud Infrastructure)

San Francisco FullTime

Posted 3mo ago

Job Location

San Francisco

Tech Stack

Remote Work Policy

On-site

Employment Type

FullTime

Categories

AI Infrastructure Engineer

About the job

The Cloud Infrastructure team is responsible for the foundational cloud primitives and deployment models that power Perplexity's products. This includes multi-tenant public cloud solutions as well as single-tenant and on-premises options for enterprise clients. As Perplexity expands its Computer and Enterprise products, this team will build and manage the security, isolation, and compliance layers essential for customer trust. They provide the deployment topologies, multi-region infrastructure, and core services that ensure Perplexity operates reliably and efficiently for both consumer traffic and large enterprises.

Responsibilities

  • Define the roadmap and technical strategy for agent-driven cloud infrastructure management.
  • Design and operate cloud networking, including VPC architectures, private connectivity, and peering for AI workloads.
  • Architect and scale compute platforms (Kubernetes/EKS, autoscaling groups, CPU/GPU fleets) for online and background workloads.
  • Build and maintain secure, isolated deployment topologies for multi-tenant, single-tenant, and BYOC environments.
  • Implement and evolve multi-region strategies for availability, failover, and data locality.
  • Collaborate with security teams to deliver enterprise controls like BYOK/KMS integrations and network isolation.
  • Develop automation, tooling, and runbooks for predictable day-2 operations across all environments.

Requirements

  • Extensive experience designing and operating cloud infrastructure on AWS (VPC design, routing, security groups, load balancing, private connectivity).
  • Strong background with Kubernetes/EKS and container orchestration, including multi-cluster, multi-region, or multi-account setups.
  • Hands-on experience with cloud networking and peering (VPC peering, Transit Gateway, private link/service endpoints, or similar).
  • Experience building or operating secure, isolated environments for enterprise customers (single-tenant, BYOC, or on-prem), including BYOK/KMS integrations and compliance.
  • Proficiency with infrastructure as code (Terraform) and strong software engineering skills in Python, Go, or Rust.
  • Strong debugging and incident management skills across distributed systems (networking, compute, platform).
  • Proven track record of driving root-cause analysis and long-term fixes.
  • 7+ years of industry experience building and operating production cloud infrastructure.
  • Experience leading the design of complex systems or migrations.

About Perplexity AI

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.