> Markdown version of [/jobs/ext/2709601-manager-web-and-mobile-site-reliability-engineering](https://www.wearedevelopers.com/jobs/ext/2709601-manager-web-and-mobile-site-reliability-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Web and Mobile Site Reliability Engineering - **Company:** Holland, Inc - **Location:** Fort Lauderdale, FL, United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Microsoft Azure, Cyber Security, Computer Engineering, DevOps, Performance Tuning, Reliability Engineering, Site Reliability Engineering Practices, Akamai, Web Platforms, Datadog, Delivery Pipeline, Grafana, Gitlab, Containerization, Kubernetes, Information Technology, Terraform, Ddos, Dynatrace, Docker - **Published:** September 4, 2026 - **Apply:** https://eicl.fa.em5.oraclecloud.com/hcmUI/CandidateExperience/en/sites/CX/job/13908/apply/email ## About the Role * Knowledge: In-depth understanding of Site Reliability Engineering practices, cloud platforms (AWS/Azure), containerization (Kubernetes, Docker), Akamai/CDN edge routing, bot detection, and CI/CD pipelines (GitLab). * Skills: Production incident management, automated infrastructure management (Terraform), performance tuning, distributed tracing, metrics-driven SLI/SLO establishment. * Abilities: Ability to lead teams during critical production outages, drive cross-functional engineering accountability for reliability, and automate operational workflows., * Bachelor's degree in Computer Science, Computer Engineering, System Administration, or equivalent experience. * 6+ years in Site Reliability Engineering, DevOps, or Infrastructure Engineering. * 2+ years of leadership or direct engineering management experience. Travel: Less than 25% with shoreside travel likely Work Conditions: Work primarily in a climate-controlled environment with minimal safety/health hazard potential. Physical Demands: Remain in a stationary position at a desk and/or computer for extended periods of time; reasonable accommodations will be offered. ## Description The Manager, Web and Mobile Site Reliability Engineering (SRE) leads the engineering team responsible for ensuring maximum uptime, high availability, performance, and resilience for enterprise web applications, mobile app backends, and public API endpoints. This role defines reliability standards, oversees 24/7 incident response, manages edge infrastructure and bot mitigation, and drives automated deployment and observability pipelines., * Team Leadership & SRE Operations: Lead and develop a high-performing team of SRE and DevOps engineers supporting 24/7 high-volume web and mobile systems. Manage on-call rotations, incident command protocols, and operational readiness. * Reliability & Observability Governance: Establish Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets. Architect end-to-end monitoring, tracing, and alerting strategies using tools like Datadog, Dynatrace, or Grafana. * Incident Management & Remediation: Lead major incident response efforts, drive blameless post-mortems, and collaborate with engineering teams to prioritize root-cause fixes and architectural resiliency improvements. * Traffic, Edge & Security Management: Partner with IT Security (PCL IT Security) and CDN providers (Akamai) to implement bot mitigation strategies, DDoS defense, WAF rules, and edge caching for key APIs and digital endpoints. * Administrative: Perform all other administrative and organizational duties as required (time keeping, training, travel, collaboration and correspondence, etc.) Knowledge & Skills: * Scope: Direct management of SRE and DevOps engineers. Operational oversight for consumer-facing web platforms, mobile backend APIs, edge routing networks, and cloud deployment pipelines. * Problem Solving: Rapidly diagnoses and mitigates complex system outages, performance bottlenecks, traffic anomalies, bot campaigns, and infrastructure failures in high-volume production environments.Resolves highly complex, enterprise-scale operational challenges that impact guest operations, maritime services, revenue-generating systems, regulatory requirements, and technology service availability. Anticipates emerging operational risks, evaluates competing business priorities, establishes governance frameworks, and makes decisions where significant operational, financial, service, and reputational consequences may exist. Develops innovative solutions to improve enterprise resilience, scalability, and operational effectiveness. * Impact: Directly ensures continuous operational availability, system security, optimal site performance, and guest trust across web and mobile touchpoints. * Leadership: The role requires strong leadership skills. Requires strong incident command leadership, strategic operational decision-making, calm under pressure, and collaborative mentorship. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [WeAreDevelopers LIVE - Chrome for Sale? Comet - the upcoming perplexity browser Stealing and leaking](https://www.wearedevelopers.com/videos/1331-wearedevelopers-live-chrome-for-sale-comet-the-upcoming-perplexity-browser-stealing-and-leaking) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Best Companies in the Netherlands: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/193-best-companies-in-the-netherlands-top-25-companies-in-2023) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Software Developer Salary in The Netherlands [2023]](https://www.wearedevelopers.com/magazine/217-software-developer-salary-in-the-netherlands-2023)