Staff Infrastructure Engineer

FanDuel Inc
Atlanta, United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$159,000.0 - $208,950.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Build Automation Cloud Computing Software Debugging Distributed Systems Network Segmentation Reliability Engineering Zero Trust Network Access Datadog Istio HybridCloud Kubernetes
+4 more
Build Tools Linkerd (Service Mesh) Software Coding Terraform

Job description

This role sits in a key spot. You’ll work directly with the Infrastructure Engineering Director to set standards for how we run Kubernetes and manage service-to-service communication. You’ll mentor engineers on what’s actually happening inside the cluster and the mesh, not just how to use them. You’ll lead incident response when things break, build automation to reduce toil, and help teams adopt patterns that work reliably at our scale.

You’re a hands-on builder who understands distributed systems at depth. You care about how things work underneath, why failures happen, and how to prevent them. You write code and build tooling. You jump into incidents. You mentor others. By doing this work well, you’ll directly improve reliability, performance, and how confident our teams feel running services on our infrastructure.

In addition to the specific responsibilities outlined above, employees may be required to perform other such duties as assigned by the Company. This ensures operational flexibility and allows the Company to meet evolving business needs.

THE GAME PLAN Everyone on our team has a part to play *

  • Design and operate our Kubernetes platform across EKS clusters. Set standards for cluster configuration, workload isolation, resource management, and cost optimization. Build the runbooks and automation that make operations predictable.
  • Own the Kong Mesh ecosystem. Mature it from a deployment into a production-grade platform with clear patterns for service-to-service security, traffic management, and observability.
  • Define and implement service-to-service communication patterns. Work with teams on zero-trust networking, certificate lifecycle, mTLS policies, and how to debug when things go wrong.
  • Build infrastructure-as-code, automation, and tooling that reduce toil and let teams operate reliably.
  • Define what good looks like for our Infrastructure platform. Set SLOs, monitor reliability, and hold the line on quality.
  • Lead incident response when Kubernetes or mesh issues impact services. Understand root cause, fix it, and make sure it doesn’t happen the same way twice.
  • Mentor engineers on Kubernetes internals, mesh design, and distributed systems thinking. Raise the bar on technical understanding across the team.
  • Evaluate infrastructure tools and platforms. Understand what to build, what to buy, and what to integrate.
  • Work with platform and product teams to understand what they need from Kubernetes and mesh. Feed that into our roadmap and direction.
  • Contribute to broader infrastructure initiatives on observability, cost optimization, and resilience. Own technical scope on assigned work. A Sneak Peek Into Our Tech Stack AWS , Kubernetes (EKS, 100+ clusters), Kong Mesh, Terraform, Helm, Datadog.

Requirements

  • 7+ years working with platform infrastructure, SRE, cloud infrastructure, or related work. You’ve built and operated large systems and shipped real things.
  • Deep hands-on experience with Kubernetes. You understand cluster architecture, multi-cluster operations, scheduling, networking, security, and how to debug when things break.
  • Strong expertise with modern service mesh platforms. Experience with Kong Mesh, Istio, Linkerd, or Envoy-based systems. Comfortable with mTLS, traffic policies, and zero-trust networking.
  • Working knowledge of AWS. You understand VPCs, networking, security, and how to operate in hybrid environments including Outposts.
  • You understand distributed systems. You know the tradeoffs involved in service-to-service communication at scale. You think about failure modes.
  • You’ve defined and tracked SLOs/SLIs for infrastructure services. You use metrics that matter to users and the business, not just technical metrics.
  • You code in at least one modern language. You build tools and automation. You’re comfortable both writing code and operating systems.
  • You’ve driven operational improvements through automation. You see toil and build solutions that scale. You prevent the same failure from happening twice.
  • You mentor and influence other engineers. You raise the technical bar. You help people understand why things work the way they do.
  • You communicate well. You can explain technical constraints to non-technical people. You influence technical direction.
  • You care about doing things right. You own problems end-to-end. You push for continuous improvement. Bonus

  • You’ve used Envoy in production.
  • You’ve operated AWS Outposts or hybrid cloud infrastructure.
  • You hold CNCF Kubernetes certifications (CKA, CKS, CKAD).
  • You’ve worked in regulated industries where network segmentation, auditability, strict controls, and uptime matter.
  • You’re familiar with Datadog or other observability platforms. You’ve used eBPF-based networking tools like Cilium. Don’t check all the boxes? That’s okay! We encourage you to still apply if you feel like you possess an adjacent skill set and are interested in learning more about this position.

Benefits & conditions

We offer amazing benefits above and beyond the basics. We have an array of health plans to choose from (some as low as $0 per paycheck) that include programs for fertility and family planning, mental health support, and fitness benefits. We offer generous paid time off (PTO & sick leave), annual bonus and long-term incentive opportunities (based on performance), 401k with up to a 5% match, commuter benefits, pet insurance, and more - check out all our benefits here: FanDuel Total Rewards. *Benefits differ across location, role, and level., The applicable salary range for this position is $159,000 - $208,950 USD, which is dependent on a variety of factors including relevant experience, location, business needs and market demand. This role may offer the following benefits: medical, vision, and dental insurance; life insurance; disability insurance; a 401(k) matching program; among other employee benefits. This role may also be eligible for short-term or long-term incentive compensation, including, but not limited to, cash bonuses and stock program participation. This role includes paid personal time off and 14 paid company holidays. FanDuel offers paid sick time in accordance with all applicable state and federal laws.

About the company

FanDuel Group is the premier mobile gaming company in the United States and Canada. FanDuel Group consists of a portfolio of leading brands across mobile wagering including: America’s #1 Sportsbook, FanDuel Sportsbook; its leading iGaming platform, FanDuel Casino; the industry’s unquestioned leader in horse racing and advance-deposit wagering, FanDuel Racing; and its daily fantasy sports product.

In addition, FanDuel Group operates FanDuel TV, its broadly distributed linear cable television network and FanDuel TV+, its leading direct-to-consumer OTT platform. FanDuel Group has a presence across all 50 states, Canada, and Puerto Rico.

The company is based in New York with US offices in Los Angeles, Atlanta, and Jersey City, as well as global offices in Canada and Scotland. The company’s affiliates have offices worldwide, including in Ireland, Portugal, Romania, and Australia.

FanDuel Group is a subsidiary of Flutter Entertainment, the world’s largest sports betting and gaming operator with a portfolio of globally recognized brands and traded on the New York Stock Exchange (NYSE: FLUT).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

1:36 min

Visualizing memory limits and isolating suspicious endpoints

Dina Matveev Dina Matveev · Europe 2026 Virtual

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

7:15 min

Installing Istio programmatically with bash scripts

Thomas Südbröcker · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all