> Markdown version of [/jobs/ext/1975470-principal-software-development-engineer-platform-infrastructure](https://www.wearedevelopers.com/jobs/ext/1975470-principal-software-development-engineer-platform-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Software Development Engineer - Platform & Infrastructure - **Company:** Expedia Inc. - **Location:** San Jose, CA, United States - **Experience:** Expert - **Salary:** $249,000.0 - $348,500.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Application Release Automation, Cloud Computing, Cloud Engineering, Code Review, Continuous Integration, Data Stores, Github, Identity and Access Management, Python (Programming Language), Key Management, Performance Tuning, Software Architecture, Cloud Services, Prometheus, Software Engineering, Spinnaker, Strategies of Testing, Datadog, Policy as Code, Cloud Platform System, Autoscaling, Istio, Grafana, Kubernetes Helm Charts, Caching, Amazon Virtual Private Cloud (VPC), Cloudformation, Amazon Relational Database Service, Containerization, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Production Code, Linkerd (Service Mesh), Api Gateway, Terraform, Jenkins, Golang, Microservices - **Published:** August 7, 2026 - **Apply:** https://www.dice.com/job-detail/a3274a91-05c4-4b43-9f57-6efcf92c3dce ## About the Role * 10+ years professional software engineering experience with significant hands-on experience building and operating distributed cloud services * Strong, demonstrable experience building and running platforms on AWS (EKS, ECS, EC2, VPC, IAM, S3, RDS, ELB/ALB, Auto Scaling) * Deep hands-on experience with containerization and Kubernetes at scale (EKS or comparable) * Practical experience with infrastructure as code (Terraform, CloudFormation) and Helm charts * Proven record of contributing production code and platform automation (languages such as Go, Python, Java, or similar) * Experience designing for resilience, observability, security and operational automation * Familiarity with CI/CD tooling and developer workflows (Spinnaker, Jenkins, GitHub Actions, GitLab CI, or similar) * Demonstrated ability to lead cross-team technical initiatives and influence architectural decisions * Strong communication skills and experience mentoring engineers, * Prior platform engineering, SRE, or infrastructure leadership at scale * Experience with monitoring/observability stacks (Prometheus, Grafana, Datadog, OpenTelemetry, Jaeger) * Proven experience with AWS cost optimization strategies and tooling (Cost Explorer, Trusted Advisor, billing APIs) * Knowledge of service meshes (Istio, Linkerd, VPC Lattice), API gateways, and advanced networking patterns * Experience with security/compliance for cloud environments, secrets management, and policy-as-code * Experience migrating monoliths to cloud-native architectures ## Description We are hiring a hands-on Principal Software Development Engineer to lead the design, implementation and operational excellence of our cloud infrastructure and platform capabilities. This role emphasizes platform engineering: designing and delivering a scalable, reliable, secure, observable, and cost-efficient runtime platform (Kubernetes, containers, CI/CD, IaC, cloud services) used by multiple product teams. You will both set strategy and be embedded in the code and infrastructure to execute the roadmap end-to-end. You will shape the foundational platform that powers customer-facing services across brands, increasing developer velocity, reducing operational risk, and optimizing cloud spend. This role is ideal for a leader who can design complex systems, implement solutions in production, and lead broad cross-team adoption. What you will do: * Define the technical strategy, roadmap and standards for cloud infrastructure and platform capabilities (container runtime, orchestration, networking, CI/CD, observability, security, IaC) * Lead migration and platform adoption efforts (containerization, Kubernetes/EKS, runtime platform) and drive roadmap execution end-to-end * Be an active code and IaC contributor (Terraform/CloudFormation/Helm, platform services, automation) and perform design/code reviews * Design scalable, resilient, and secure infrastructure patterns for microservices, data stores, caching, and messaging * Build and improve CI/CD pipelines, release automation, testing strategies, and safe deployment practices * Drive SRE and observability practices: define SLIs/SLOs, monitoring, tracing, alerting, runbooks, incident response, and post-mortems * Lead capacity planning, performance tuning, and traffic engineering for high-scale workloads * Own and evangelize AWS cost optimization practices (rightsizing, reserved/spot strategies, architecture changes, cost monitoring and show back) * Define and enforce platform standards, security controls, environment tagging, and operational excellence practices across teams * Mentor and grow senior and mid-level engineers; identify high-potential talent and raise engineering standards * Collaborate closely with product, security, platform, and operations stakeholders to align technical solutions with business goals and compliance requirements ## Related Videos - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [Why Attend a Developer Event in 2026?](https://www.wearedevelopers.com/magazine/688-why-attend-a-developer-event-in-2026)