> Markdown version of [/jobs/ext/2611972-platform-infrastructure-sre-software-engineer](https://www.wearedevelopers.com/jobs/ext/2611972-platform-infrastructure-sre-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Platform Infrastructure SRE / Software Engineer - **Company:** Bayside Solutions - **Location:** Cupertino, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $124,800.0 - $145,600.0 - **Contract:** Permanent contract - **Skills:** User Authentication, Cloud Computing, Cloud Computing Security, Digital Architecture, Distributed Data Store, Distributed Systems, Domain Name System (DNS), Routing, Reliability Engineering, Data Logging, Pulumi, Google Cloud, Load Balancing, Istio, Apache Spark, Multi-Cloud, Amazon Virtual Private Cloud (VPC), Cloudformation, Kubernetes, Infrastructure Automation Frameworks, Apache Flink, Deployment Automation, Terraform - **Published:** August 15, 2026 - **Apply:** https://www.dice.com/job-detail/d2047464-9af7-40b1-9048-0e2f3bfc7c6d ## About the Role * Deep experience with Crossplane for infrastructure provisioning and platform automation. * Experience with Alibaba Cloud and its Kubernetes, networking, and infrastructure services. * Experience with AWS EKS and/or Google Cloud Platform. * Experience designing or operating multi-cloud platforms. * Experience with service mesh technologies and Kubernetes service networking. * Experience with distributed data technologies such as Apache Spark, Apache Flink, or Trino. * Experience implementing authentication, authorization, cloud security, and governance controls. * Experience building and maintaining CI/CD pipelines for Kubernetes-based platforms. * Experience designing highly available and resilient platform architectures. * Experience automating provisioning and lifecycle management across a large number of environments. * Strong production SRE, incident response, reliability engineering, and operational automation experience. ## Description We are looking for a strong Senior Platform Infrastructure SRE / Software Engineer to support the productionization and operation of large-scale Kubernetes-based platform services across multiple cloud environments. This role is best suited for someone with deep Kubernetes and infrastructure experience who can take capabilities developed by a platform engineering team and make them repeatable, scalable, observable, reliable, and production-ready across many environments., * Productionize Kubernetes-based platform services developed by platform engineering teams. * Deploy and operate platform infrastructure across multiple cloud and production environments. * Build repeatable environment provisioning and deployment automation. * Develop reusable infrastructure templates, blueprints, and deployment patterns. * Provision Kubernetes clusters, cloud infrastructure, networking, and service dependencies. * Configure environment-specific infrastructure, connectivity, and platform services. * Deploy, validate, upgrade, and maintain platform services throughout their lifecycle. * Build monitoring, alerting, dashboards, logging, health checks, and operational controls. * Establish and validate production-readiness standards for new platform capabilities. * Troubleshoot complex failures across applications, Kubernetes, infrastructure, networking, and distributed systems. * Work closely with platform developers to understand application behavior and identify operational gaps before production rollout. * Implement and maintain Infrastructure as Code using technologies such as Crossplane, Terraform, Pulumi, or CloudFormation. * Build and maintain Helm-based Kubernetes packaging and deployment patterns. * Design safe rollout, rollback, upgrade, recovery, and lifecycle-management processes. * Validate platform capacity, availability, scalability, and reliability. * Configure and troubleshoot DNS, load balancing, VPC networking, routing, service connectivity, security policies, and certificates. * Improve operational automation and reduce manual environment-specific work. * Document architecture, deployment patterns, operational procedures, and troubleshooting guidance. * Participate in production support, incident response, and root-cause analysis as appropriate. * Independently own technical work and drive complex problems through resolution with limited supervision. ## Related Videos - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [Flex your Energy: Building a Cloud-Native Platform for Renewable Energy Communities](https://www.wearedevelopers.com/videos/1990-flex-your-energy-building-a-cloud-native-platform-for-renewable-energy-communities) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere)