> Markdown version of [/jobs/ext/3676542-staff-software-engineer-devops](https://www.wearedevelopers.com/jobs/ext/3676542-staff-software-engineer-devops). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Software Engineer, Devops - **Company:** TEAM Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $155,854.0 - $233,781.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Computing, Continuous Integration, Data as a Services, Information Engineering, Software Design Documents, DevOps, Domain Name System (DNS), Github, Identity and Access Management, Virtual Private Networks (VPN), Python (Programming Language), Key Management, Machine Learning, Routing, Octopus Deploy, Role-Based Access Control, Reliability Engineering, Prometheus, TypeScript, Virtualization Technology, Policy as Code, AWS Cdk, Enterprise Software Applications, Load Balancing, Istio, Delivery Pipeline, Grafana, Amazon Virtual Private Cloud (VPC), Backend, Kubernetes, Information Technology, Rancher, Github Enterprise, Bare Metal, Cloudwatch, Data Pipelines, Serverless Computing, Vmware - **Published:** October 10, 2026 - **Apply:** https://www.dice.com/job-detail/3278a53d-fc9d-4fdb-a154-ec9f850aba6a ## About the Role * 10+ years in DevOps, site reliability, platform or infrastructure engineering, including several years setting technical direction for production platforms. * Deep AWS Expertise: Extensive hands-on experience designing and operating production AWS environments across many accounts, including multi-account governance, identity and access management, VPC networking and hybrid connectivity, DNS, EKS, containers, serverless and managed data services. An AWS Professional or Specialty certification is a plus. * Kubernetes in Production, in the Cloud and On-Premises: Proven experience running Kubernetes in production on EKS and on self-managed clusters built on VMs or bare metal, including cluster lifecycle, CNI networking, ingress, storage, RBAC and upgrades. Experience with Rancher or a comparable multi-cluster management platform is a strong plus. * Python and Infrastructure as Code: Strong experience with infrastructure as code; AWS CDK is strongly preferred. Our platform code is written in Python, so strong Python skills are required. You can read, understand and make targeted changes to TypeScript, which some of our application teams use. * CI/CD and GitOps: Experience building delivery pipelines with GitHub Actions or similar tools, and deploying to Kubernetes with GitOps tooling (Argo CD or Flux), Helm and Kustomize. * Reliability Engineering: A track record of establishing SLOs, observability (Prometheus, Grafana, OpenTelemetry, CloudWatch or similar), incident management and on-call practices, and of participating in on-call yourself. * Hybrid Networking: A solid grasp of networking across cloud and on-premises environments: routing, VPN, DNS, load balancing, certificates and firewalls, and how each of them fails. * Technical Leadership Without Authority: Ability to drive architectural decisions across teams you don't manage, write clear design documents, and bring skeptical stakeholders along. * Startup Orientation: Comfort in a fast-moving, resource-constrained environment. You know how to ship, how to cut scope without cutting corners, and how to improve a system while it's running. * Bachelor's degree in Computer Science, Engineering or a related field, or equivalent practical experience. Nice to have * Exposure to manufacturing, industrial or OT environments and the constraints of running software near the plant floor. * Experience running GPU, machine learning or AI and agentic workloads on Kubernetes. * Familiarity with enterprise virtualization platforms such as VMware. * Policy as code (OPA Gatekeeper or Kyverno), service mesh, or GitHub Enterprise administration. WORK AUTHORIZATION REQUIREMENTApplicants must be authorized to work in the United States on a permanent basis. We are unable to offer visa sponsorship at this time. ## Description * Set the Direction for Our AWS Foundation: Evolve our multi-account AWS environment, including account governance and guardrails, identity and access management, and the networking and DNS that connect the cloud to our sites. Decide where new workloads live and how they connect. * Build the Reliability Practice: Establish our on-call rotation, SLOs and error budgets, alerting standards, incident response and blameless postmortems. Stand up the observability (metrics, logs, traces and synthetic checks) that tells us something is wrong before our users do. * Enable the Teams We Serve: Partner with our Digital teams (customer-facing website and backend services), Data Engineering, Manufacturing Systems and enterprise application teams. Give them well-supported paths: CI/CD templates, deployment patterns, reusable infrastructure libraries and self-service environments. * Plan for What's Next: Shape how the platform supports manufacturing and plant-floor workloads, data pipelines, and emerging AI and agentic workloads, whether each one runs on-premises or in the cloud. * Build In Security and Cost Discipline: Make least-privilege access, policy guardrails, secrets management and software supply chain controls the default. Keep cloud spend visible and deliberate through tagging, right-sizing and commitment planning. * Ship and Mentor: Write production infrastructure and code every week. Lead design reviews, mentor engineers on the team and across the organization, and write the documentation and runbooks that let others move without waiting on you., At Slate, we're fueled by grit, determination, and attention to detail. The start-up spirit of ingenuity and resourcefulness move our business forward. Team Slate fosters a culture of excellence, innovation, and mutual respect, and is motivated by shared principles. * Safety First * Delight Customers * One Team * Relentless Improvement * Fast, Frugal, and Scrappy * Respectful Collaboration * Positive Legacy