> Markdown version of [/jobs/ext/2698082-cloud-platform-architect](https://www.wearedevelopers.com/jobs/ext/2698082-cloud-platform-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloud Platform Architect - **Company:** Greenhouse Software, Inc. - **Location:** San Jose, CA, United States - **Experience:** Expert - **Salary:** $245,000.0 - $325,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Software as a Service, Cloud Computing, Data Centers, DevOps, Domain Name System (DNS), Github, Hypertext Transfer Protocols (HTTP), Identity and Access Management, Python (Programming Language), Networking Basics, Octopus Deploy, Open Source Technology, Reliability Engineering, Software Engineering, TCP/IP, Data Logging, Google Cloud, Load Balancing, Cloud Platform System, Istio, Multi-Cloud, Reliability of Systems, Gitlab-ci, Kubernetes, Linkerd (Service Mesh), Hardware Infrastructure, Terraform, Jenkins, Programming Languages - **Published:** September 3, 2026 - **Apply:** https://boards.greenhouse.io/embed/job_app?for=sambanovasystems&token=5719049004 ## About the Role * 7+ years of experience in DevOps, Site Reliability Engineering (SRE), or Cloud Infrastructure roles * Proficiency in at least one programming language (e.g., Python, Go, Rust) * Expertise with Kubernetes (EKS, GKE, or self-managed) in production environments - pods, operators, CRDs, CNIs, etc. * Expertise with Infrastructure as Code with the ability to manage complex, multi-cloud environments * Strong proficiency with at least one major cloud provider (AWS, GCP, or Azure), with a solid understanding of the core services (compute, storage, networking, IAM) * Networking fundamentals (TCP/IP, DNS, HTTP, load balancing) and security best practices in the cloud, * Experience in a hybrid environment bridging cloud and on-premise/data center infrastructure * Experience managing infrastructure for data-intensive or ML/AI workloads * Knowledge of building and maintaining CI/CD pipelines (e.g., GitLab CI, Jenkins, ArgoCD) * Experience with service mesh technologies (e.g., Istio, Linkerd) * Contributions to open-source projects or a public portfolio of code (GitHub) ## Description The Cloud Operations team is seeking an experienced engineering leader to scale the platform our internal and external customers use to access SambaNova RDUs., In this role, you'll architecting our next-generation system from the ground up, running the Kubernetes infrastructure that powers some of the most advanced AI workloads in the industry, and bridging multi-cloud and on-prem environments in ways no generic SaaS company can offer. Your work will directly impact the productivity of every engineer at SambaNova and by extension, the speed at which we ship the future of AI computing. Some of your responsibilities include: * Architect, build, and maintain our next-generation internal developer platform, automating and streamlining our cloud and on-prem infrastructure * Design, write, and manage Terraform modules to provision and manage resources across AWS, GCP, and Azure, ensuring consistency and reproducibility * Build and manage highly available, secure, and performant Kubernetes clusters that serve as the primary runtime for our diverse AI workloads * Design and implement robust networking solutions (VPCs, load balancers, firewalls, service meshes) that seamlessly connect our multi-cloud and hybrid environments * Collaborate with AI and software engineering teams to understand their needs, provide golden paths to production, and build internal tools that accelerate their development cycles * Implement best practices for observability (monitoring, logging, tracing) to ensure system reliability and performance, and participate in on-call rotation ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) - [Get started with securing your cloud-native Java microservices applications](https://www.wearedevelopers.com/videos/123-get-started-with-securing-your-cloud-native-java-microservices-applications) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [A Guide to Green Tech and Green IT Careers](https://www.wearedevelopers.com/magazine/374-a-guide-to-green-tech-and-green-it-careers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)