> Markdown version of [/jobs/ext/917570-member-of-technical-staff-software-engineer-cloud-infrastructure](https://www.wearedevelopers.com/jobs/ext/917570-member-of-technical-staff-software-engineer-cloud-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Member of Technical Staff (Software Engineer, Cloud Infrastructure) - **Company:** Perplexity AI - **Location:** New York, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Computing, Disaster Recovery, Distributed Systems, Python (Programming Language), Routing, Peering, Runbook, Software Engineering, Load Balancing, Amazon Virtual Private Cloud (VPC), Kubernetes, Low Latency, Terraform - **Published:** June 6, 2026 - **Apply:** https://www.dice.com/job-detail/8c7c5ffc-5bd1-4c5d-959f-6629e0ff8115 ## About the Role * Deep experience designing and operating cloud infrastructure on AWS (VPC design, routing, security groups, load balancing, private connectivity). * Strong background with Kubernetes/EKS and container orchestration, including multi-cluster, multi-region, or multi-account setups. * Hands-on experience with cloud networking and peering (VPC peering, Transit Gateway, private link/service endpoints, or similar constructs in neoclouds). * Experience building or operating secure, isolated environments for enterprise customers (single-tenant, BYOC, or on-prem), ideally including BYOK/KMS integrations and compliance constraints. * Proficiency with infrastructure as code (Terraform) and strong software engineering skills in at least one of Python, Go, or Rust for automation and tooling. * Strong debugging and incident management skills across distributed systems (networking, compute, and platform layers), with a track record of driving root-cause analysis and long-term fixes. * 7+ years of industry experience building and operating production cloud infrastructure, including leading the design of complex systems or migrations. ## Description The Cloud Infrastructure team owns the foundational cloud primitives and deployment models that power Perplexity's products, from multi-tenant public cloud to single-tenant and on-premises solutions for enterprise customers., * Own the roadmap and technical strategy for agent-driven cloud infrastructure management. * Design and operate Perplexity's cloud networking fabric, including VPC architectures, private connectivity, and peering with hyperscalers and neocloud providers to support low-latency, high-throughput AI workloads. * Architect and scale compute platforms (Kubernetes/EKS, autoscaling groups, and mixed CPU/GPU fleets) to efficiently serve online request traffic and background workloads across regions. * Build and maintain secure, isolated deployment topologies for multi-tenant, single-tenant, and customer-owned cloud (BYOC) environments, including cross-account networking, identity, and policy guardrails. * Implement and evolve multi-region strategies for availability, failover, and data locality, including traffic routing, regional capacity planning, and disaster recovery playbooks. * Partner with security to deliver enterprise controls such as BYOK/KMS integrations, network isolation, and auditability required for regulated customers. * Develop automation, tooling, and runbooks that make day-2 operations (provisioning, upgrades, incident response) predictable and repeatable for Perplexity products across all environments. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Technical Documentation - How Can I Write Them Better and Why Should I Care?](https://www.wearedevelopers.com/videos/681-technical-documentation-how-can-i-write-them-better-and-why-should-i-care) - [Transforming Education: A Journey from interactive Markdown to Remote-Labs](https://www.wearedevelopers.com/videos/941-transforming-education-a-journey-from-interactive-markdown-to-remote-labs) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Bridging AI and Nomad: a Go-based MCP Server for Cluster Control](https://www.wearedevelopers.com/videos/2063-bridging-ai-and-nomad-a-go-based-mcp-server-for-cluster-control) - [P2P networks in Blockchain](https://www.wearedevelopers.com/videos/209-p2p-networks-in-blockchain) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)