Slalom Flex (Project Based) - Site Reliability Engineer

Slalom, LLC
Atlanta, GA, United States
15 days ago
Apply on www.jofdav.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$145,600.0 - $176,800.0
Working hours
Regular working hours
Job source

Tech stack

Disaster Recovery Distributed Systems Fault Tolerance Performance Tuning Reliability Engineering Software Engineering

Job description

  • Define and establish enterprise reliability standards, Site Reliability Engineering (SRE) practices, SLIs/SLOs, operational governance models, and engineering guardrails that enable scalable, resilient technology platforms.
  • Lead the design and implementation of observability, monitoring, alerting, incident management, and operational excellence frameworks that improve visibility, accelerate issue resolution, and drive continuous service improvement.
  • Partner closely with product, engineering, infrastructure, and business leaders to embed reliability, resilience, and operational readiness into the software development lifecycle.
  • Drive enterprise resilience initiatives including capacity planning, performance engineering, disaster recovery, business continuity, and fault-tolerance strategies.
  • Establish reliability metrics, operational processes, governance mechanisms, and accountability models that reduce operational risk and improve system health throughout modernization and transformation efforts.
  • Navigate complex organizational environments, aligning cross-functional stakeholders and driving consensus on reliability priorities, standards, and operating models.
  • Translate ambiguous business problems and evolving technology needs into clear strategies, scalable solutions, and measurable outcomes.
  • Enable product and engineering teams to innovate and move quickly while implementing the operational guardrails necessary for enterprise scale, security, governance, and reliability.
  • Serve as a trusted advisor to technology leadership, influencing strategic decisions related to platform reliability, operational maturity, and organizational readiness.

Requirements

  • Deep experience building and scaling enterprise reliability, operational governance, observability, or SRE programs within complex technology organizations.
  • Proven success defining standards, policies, operating models, and governance frameworks from the ground up.
  • Strong understanding of reliability engineering concepts, including SLIs, SLOs, incident management, capacity planning, disaster recovery, performance optimization, and operational excellence practices.
  • Experience designing observability and monitoring strategies that provide actionable insights across distributed systems and modern cloud environments.
  • Demonstrated ability to operate effectively in highly ambiguous environments, transforming undefined concepts into practical, scalable solutions.
  • Exceptional stakeholder management and influencing skills, with a track record of aligning business, product, and technology leaders around a shared vision.
  • Ability to navigate organizational complexity and competing priorities while driving outcomes across multiple teams and functions.
  • Strategic mindset paired with a hands-on approach, capable of moving seamlessly between executive-level planning and execution details.
  • Strong communication skills with the ability to articulate technical concepts to both technical and non-technical audiences.
  • Passion for enabling engineering velocity while establishing the governance, risk management, security, and operational standards required for enterprise-scale growth.

Benefits & conditions

Slalom prides itself on helping team members thrive in their work and life. As a result, Slalom is proud to invest in benefits that includemeaningful time off and paid holidays, 401(k) with a match, a range of choices for highly subsidized health, dental, & vision coverage, adoption and fertility assistance, and short/long-term disability. We also offer yearly $350 reimbursement account for any well-being-related expenses, as well as discounted home, auto, and pet insurance.

Slalom is committed to fair and equitable compensation practices. For this position, the pay range is $70/hr to $85/hr. Actual compensation will depend upon an individual’s skills, experience, qualifications, location, and other relevant factors. The salary pay range is subject to change and may be modified at any time.

About the company

At Slalom, personal connection meets global scale. Our vision is to enable a world in which everyone loves their work and life. We help organizations of all kinds redefine what’s possible, give shape to the future-and get there., Slalom is a fiercely human business and technology consulting company that leads with outcomes to bring more value, in all ways, always. From strategy through delivery, our agile teams across 52 offices in 12 countries collaborate with clients to bring powerful customer experiences, innovative ways of working, and new products and services to life. We are trusted by leaders across the Global 1000, many successful enterprise and mid-market companies, and 500+ public sector organizations to improve operations, drive growth, and create value. At Slalom, we believe that together, we can move faster, dream bigger, and build better tomorrows for all.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jofdav.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:05 min

Practical Byzantine Fault Tolerance in distributed computing systems

Jonan Scheffler · World Congress 2022

2:14 min

Crafting an effective disaster recovery and communication plan

Mihaela-Roxana Ghidersa · LIVE

8:32 min

Benchmarking GitOps engine constraints for extensive multi-cluster environments

Artem Lajko · Europe 2026 Virtual

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

3:49 min

Enhancing system resilience and fault tolerance

Michael Eder +1 · LIVE

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

Videos

See all

Related articles

See all