> Markdown version of [/jobs/ext/3102632-staff-devops-engineer-on-platform-infrastructure](https://www.wearedevelopers.com/jobs/ext/3102632-staff-devops-engineer-on-platform-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff DevOps Engineer on Platform Infrastructure - **Company:** Red Cell Partners - **Location:** Seattle, WA, United States - **Experience:** Expert - **Salary:** $180,000.0 - $240,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Applicant Tracking Systems, Application Packaging, Application Services, Microsoft Azure, Software as a Service, Databases, Continuous Integration, Data Retention, Linux, DevOps, Distributed Systems, Domain Name System (DNS), Fraud Prevention and Detection, Identity and Access Management, Python (Programming Language), Key Management, Software Tools, Cloud Services, Software Engineering, TypeScript, Private Cloud Environment, Pulumi, Google Cloud, Load Balancing, Istio, HybridCloud, Infrastructure as Code (IaC), Containerization, Kubernetes, Infrastructure Automation Frameworks, Hardware Infrastructure, Terraform, Docker, Golang - **Published:** September 27, 2026 - **Apply:** https://www.dice.com/job-detail/b9e3d9aa-2b22-4531-9ebe-87e56f0eee64 ## About the Role * 10+ years of software, platform, infrastructure, SRE, or DevOps engineering experience, including ownership of production systems. * Deep hands-on experience with Linux, Docker, Kubernetes, and production cluster operations. * Strong experience with Helm and infrastructure as code, such as Terraform or Pulumi, including reusable modules, state management, testing, and change review. * Hands-on infrastructure experience with at least two of AWS, Azure, and Google Cloud Platform, with working knowledge of core compute, networking, storage, identity, and managed-service patterns across all three. * Experience deploying and operating software across customer-controlled, private-cloud, hybrid-cloud, or on-premises environments, including adapting cloud-native SaaS products for these deployment models. * Strong understanding of networking, DNS, ingress, load balancing, certificates, IAM, secrets management, persistent storage, and service-to-service security. * Experience building and operating CI/CD or GitOps systems, including production releases, upgrades, and rollbacks, with metrics, logs, traces, SLOs, and alerts for incident response. * Strong software engineering and automation skills in Python, Go, TypeScript, or a similar language, with the ability to work across application and infrastructure layers. * Experience designing secure, portable production infrastructure, including hardening systems and replacing or abstracting cloud-specific managed services when needed. * Experience converting a cloud-native SaaS product into a customer-hosted, private-cloud, or on-premises deployment model. * Demonstrated experience using AI-assisted coding and engineering tools to accelerate development, infrastructure automation, troubleshooting, operational analysis, or incident investigation. Preferred * Experience operating in restricted-network, disconnected, regulated, or security-sensitive environments. * Familiarity with compliance and security frameworks such as HIPAA, SOC 2, NIST, FedRAMP, or related government requirements. * Experience with service mesh, policy as code, admission controls, software supply-chain security, artifact signing, or software bills of materials. * Hands-on experience with Crossplane or meaningful contributions to CNCF projects, such as code, documentation, design, or community maintenance. * Experience supporting long-running, stateful, data-intensive, AI/ML, or GPU-enabled workloads. * Customer-facing engineering, solutions architecture, or forward-deployed engineering experience. ## Description As a Senior or Staff DevOps Engineer on Platform Infrastructure, you will design and operate the infrastructure that supports Trase OS across Trase-hosted and customer-controlled environments. You will build secure, repeatable deployment patterns spanning AWS, Microsoft Azure, Google Cloud Platform, private cloud, hybrid cloud, and on-premises infrastructure. This is a hands-on engineering role with broad ownership. You will work across application packaging, Kubernetes, infrastructure as code (IaC), networking, security, observability, release engineering, and production reliability. You will also partner directly with engineering and customer-facing teams to turn deployment requirements into systems that can be installed, upgraded, operated, and supported consistently. The level will reflect your experience and demonstrated scope. Staff-level candidates will be expected to lead architecture across teams, establish engineering standards, and mentor other engineers. Why this Role is Needed Trase OS supports mission-critical, long-running workflows in environments with different cloud services, network controls, security requirements, and operating models. Our deployment architecture must remain portable without sacrificing reliability, security, or operational clarity. This role will reduce one-off deployment work, remove avoidable dependencies on a single cloud provider, and establish reusable infrastructure that internal teams and customers can operate with confidence. What You'll Do * Architect, build, and operate secure infrastructure across AWS, Microsoft Azure, and Google Cloud Platform, as well as private-cloud, hybrid-cloud, on-premises, and customer-controlled environments. * Containerize and package Trase OS application services using Docker, Kubernetes, Helm, Kustomize, or equivalent tools. * Create reusable infrastructure-as-code (IaC) modules and deployment workflows using Terraform, Pulumi, or comparable tooling. * Design deployment patterns that account for customer-specific requirements such as restricted networks, limited or no egress, approved registries, data residency, and cloud account ownership. * Identify, replace, or abstract hard dependencies on managed cloud services when they prevent portability across deployment environments. * Troubleshoot complex issues across applications, Kubernetes clusters, cloud services, networks, and infrastructure rather than treating platform work as CI/CD scripting alone. * Build and maintain CI/CD and GitOps workflows, release orchestration, environment promotion, upgrade paths, rollback procedures, and version compatibility controls. * Build and operate observability for Trase OS across metrics, logs, traces, dashboards, alerting, and SLOs so teams can diagnose failures and support customer deployments. Technical Leadership * Lead infrastructure and reliability decisions that affect multiple product and customer teams. * Set practical standards for production readiness, cloud portability, security, observability, and operational support. * Make clear tradeoffs between speed, maintainability, cost, security, and customer requirements, and drive decisions through implementation. * Build reusable platform capabilities that replace one-off deployment solutions. * Mentor engineers and raise the team's ability to design, ship, and operate distributed systems., By submitting an application, you acknowledge that Red Cell Partners, LLC ("Red Cell") uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, "Hiring Platforms"). Your application materials, including your rsum, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes: * Managing and administering your application throughout the hiring process; * Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases; * Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories. Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact Red Cell requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Red Cell's data retention policies. ## Related Videos - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Scoring 2000 Products per Request: Performance Pitfalls in Golang](https://www.wearedevelopers.com/videos/2073-scoring-2000-products-per-request-performance-pitfalls-in-golang) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)