> Markdown version of [/jobs/ext/2714868-staff-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2714868-staff-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Infrastructure Engineer - **Company:** FanDuel Inc - **Location:** Atlanta, United States - **Experience:** Expert - **Salary:** $159,000.0 - $208,950.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Build Automation, Cloud Computing, Software Debugging, Distributed Systems, Network Segmentation, Reliability Engineering, Zero Trust Network Access, Datadog, Istio, HybridCloud, Kubernetes, Build Tools, Linkerd (Service Mesh), Software Coding, Terraform - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/staff-infrastructure-engineer-fanduel-8581803 ## About the Role + 7+ years working with platform infrastructure, SRE, cloud infrastructure, or related work. You've built and operated large systems and shipped real things. + Deep hands-on experience with Kubernetes. You understand cluster architecture, multi-cluster operations, scheduling, networking, security, and how to debug when things break. + Strong expertise with modern service mesh platforms. Experience with Kong Mesh, Istio, Linkerd, or Envoy-based systems. Comfortable with mTLS, traffic policies, and zero-trust networking. + Working knowledge of AWS. You understand VPCs, networking, security, and how to operate in hybrid environments including Outposts. + You understand distributed systems. You know the tradeoffs involved in service-to-service communication at scale. You think about failure modes. + You've defined and tracked SLOs/SLIs for infrastructure services. You use metrics that matter to users and the business, not just technical metrics. + You code in at least one modern language. You build tools and automation. You're comfortable both writing code and operating systems. + You've driven operational improvements through automation. You see toil and build solutions that scale. You prevent the same failure from happening twice. + You mentor and influence other engineers. You raise the technical bar. You help people understand why things work the way they do. + You communicate well. You can explain technical constraints to non-technical people. You influence technical direction. + You care about doing things right. You own problems end-to-end. You push for continuous improvement. Bonus + You've used Envoy in production. + You've operated AWS Outposts or hybrid cloud infrastructure. + You hold CNCF Kubernetes certifications (CKA, CKS, CKAD). + You've worked in regulated industries where network segmentation, auditability, strict controls, and uptime matter. + You're familiar with Datadog or other observability platforms. You've used eBPF-based networking tools like Cilium. Don't check all the boxes? That's okay! We encourage you to still apply if you feel like you possess an adjacent skill set and are interested in learning more about this position. ## Description This role sits in a key spot. You'll work directly with the Infrastructure Engineering Director to set standards for how we run Kubernetes and manage service-to-service communication. You'll mentor engineers on what's actually happening inside the cluster and the mesh, not just how to use them. You'll lead incident response when things break, build automation to reduce toil, and help teams adopt patterns that work reliably at our scale. You're a hands-on builder who understands distributed systems at depth. You care about how things work underneath, why failures happen, and how to prevent them. You write code and build tooling. You jump into incidents. You mentor others. By doing this work well, you'll directly improve reliability, performance, and how confident our teams feel running services on our infrastructure. In addition to the specific responsibilities outlined above, employees may be required to perform other such duties as assigned by the Company. This ensures operational flexibility and allows the Company to meet evolving business needs. THE GAME PLAN Everyone on our team has a part to play * + Design and operate our Kubernetes platform across EKS clusters. Set standards for cluster configuration, workload isolation, resource management, and cost optimization. Build the runbooks and automation that make operations predictable. + Own the Kong Mesh ecosystem. Mature it from a deployment into a production-grade platform with clear patterns for service-to-service security, traffic management, and observability. + Define and implement service-to-service communication patterns. Work with teams on zero-trust networking, certificate lifecycle, mTLS policies, and how to debug when things go wrong. + Build infrastructure-as-code, automation, and tooling that reduce toil and let teams operate reliably. + Define what good looks like for our Infrastructure platform. Set SLOs, monitor reliability, and hold the line on quality. + Lead incident response when Kubernetes or mesh issues impact services. Understand root cause, fix it, and make sure it doesn't happen the same way twice. + Mentor engineers on Kubernetes internals, mesh design, and distributed systems thinking. Raise the bar on technical understanding across the team. + Evaluate infrastructure tools and platforms. Understand what to build, what to buy, and what to integrate. + Work with platform and product teams to understand what they need from Kubernetes and mesh. Feed that into our roadmap and direction. + Contribute to broader infrastructure initiatives on observability, cost optimization, and resilience. Own technical scope on assigned work. A Sneak Peek Into Our Tech Stack AWS , Kubernetes (EKS, 100+ clusters), Kong Mesh, Terraform, Helm, Datadog. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Building a Cloud Platform Where Everything is Just Another Kubernetes Resource](https://www.wearedevelopers.com/videos/100137-building-a-cloud-platform-where-everything-is-just-another-kubernetes-resource) - [Implementing Feature Environments with AWS and Terraform](https://www.wearedevelopers.com/videos/531-implementing-feature-environments-with-aws-and-terraform) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)