Director, Infrastructure & IT Operations

Peoplefinders, LLC
United States
4 days ago
Apply on www.builtincolorado.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$170,000.0 - $180,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Backup Devices Cloud Computing Cloud Engineering Configuration Management Collaborative Software Databases Continuous Integration DevOps Disaster Recovery
+24 more
Distributed Systems Identity and Access Management Information Technology Operations Productivity Software Reliability Engineering Cloud Services User Provisioning Software Software Vulnerability Management Web Applications EndPointSecurity Data Logging Cloud Platform System Captcha Rate Limiting Containerization Infrastructure Automation Frameworks Information Technology Cloudflare Search Engines Laptops BIG-IP Access Policy Manager (APM) Appdynamics Dynatrace Unified Endpoint Management

Job description

We are seeking a strategic and hands-on Director, Infrastructure & IT Operations to lead the technology capabilities that keep our digital products, cloud platforms, and employees secure, reliable, and productive.

This leader will oversee cloud infrastructure, Site Reliability Engineering, production operations, bot mitigation, corporate IT, and helpdesk support, while providing architectural leadership across cloud platforms, networking, identity, observability, and operational tooling.

The ideal candidate combines strong technical judgment with disciplined operational leadership: building high-performing teams, setting clear service expectations, improving reliability, managing vendors and costs, and reducing operational risk across the company’s portfolio of digital properties.

This is not a traditional corporate IT position - it spans both internal employee technology and the production infrastructure supporting high-traffic, customer-facing applications.

Key Responsibilities

Infrastructure and Technology Leadership

Lead the teams responsible for cloud infrastructure, SRE, production operations, corporate IT, helpdesk, and bot mitigation. Establish ownership, operational standards, SLOs, escalation procedures, and performance metrics for each function. Develop managers, technical leads, and IT staff through coaching and career development. Build quarterly and annual infrastructure roadmaps aligned with business priorities, growth, security, reliability, and cost. Partner with Engineering, Product, Data, Marketing, Finance, HR, Legal, and executives, providing clear recommendations on risks, architecture, and tradeoffs. Build a culture of accountability, documentation, automation, and continuous improvement.

Cloud Infrastructure and Architecture

Own the reliability, scalability, security, performance, and cost effectiveness of the company’s cloud infrastructure, primarily AWS. Provide architectural leadership across cloud services, applications, APIs, networking, identity, databases, storage, observability, and integrations. Establish architecture standards, reference patterns, governance, and technical review processes. Partner with engineering to improve deployment safety, performance, and readiness. Drive automation, configuration management, and infrastructure as code. Lead capacity planning for business growth and demand spikes. Eliminate single points of failure, undocumented systems, manual processes, and reliance on individual employees or vendors. Review new systems for architectural fit, supportability, security, and total cost of ownership. Ensure production and nonproduction environments are properly separated, secured, and monitored.

Site Reliability and Production Operations

Establish and maintain SLIs, SLOs, availability targets, and error-budget practices for critical systems. Improve observability through centralized logging, metrics, distributed tracing, APM, and actionable alerting. Lead major incident management, including coordination, executive communication, root-cause analysis, and corrective-action tracking. Develop and test backup, disaster recovery, and business continuity procedures. Improve mean time to detect, respond, and recover from incidents. Establish on-call practices balancing coverage with team sustainability. Partner with development teams to improve resilience and production support. Use incident data to prioritize long-term reliability work. Ensure systems have current runbooks, architecture diagrams, and documentation.

Bot Mitigation and Traffic Protection

Own the strategy for bot mitigation, scraping protection, automated abuse prevention, and traffic-quality management. Protect company properties from scraping, credential attacks, fraud, automated abuse, and infrastructure exhaustion. Lead configuration of technologies such as Cloudflare, DataDome, WAFs, rate limiting, CAPTCHA alternatives, device fingerprinting, and application-level defenses. Partner with Engineering, Product, Marketing, SEO, and Revenue to distinguish malicious automation from legitimate customers, search engines, partners, and approved AI crawlers. Establish monitoring and response procedures for scraping and abnormal traffic. Measure effectiveness through traffic reduction, false-positive rates, platform stability, and infrastructure savings. Evaluate vendors for measurable value, and stay current on scraping techniques, browser automation, residential proxies, and AI crawlers.

Corporate IT and Helpdesk

Oversee internal IT services and helpdesk support for employees. Establish service-level expectations for response time, resolution time, satisfaction, and backlog reduction. Own the employee technology lifecycle: onboarding, offboarding, equipment provisioning, access management, and asset recovery. Manage laptops, endpoints, collaboration platforms, productivity software, and telecom services. Standardize endpoint configuration, encryption, patching, and endpoint detection. Improve the support experience through clear intake, self-service resources, and automation. Maintain accurate inventories of devices, licenses, and assets. Partner with HR so new employees get equipment and access on time and departing employees have access promptly removed. Identify and address the root causes of recurring support issues.

Identity, Access, and Operational Security

Establish consistent identity and access-management practices across corporate and production systems. Implement role-based access, least-privilege principles, MFA, periodic access reviews, and timely access removal. Maintain secure processes for privileged, administrative, and service accounts. Partner with security and legal to identify and reduce technology risks. Support vulnerability remediation, security assessments, audits, and vendor reviews. Ensure cloud environments, endpoints, applications, and third-party services meet company security standards, and maintain readiness for cybersecurity incidents and applicable privacy requirements.

Vendor, Budget, and Cost Management

Manage infrastructure, security, IT, telecom, monitoring, and support vendors, including evaluations, contract reviews, renewals, licensing, and SLAs. Develop and manage infrastructure and IT budgets. Improve cloud cost visibility through tagging, allocation, and forecasting. Identify opportunities to eliminate unused services, consolidate tools, and renegotiate contracts. Communicate budget performance and investment recommendations to leadership. Ensure critical vendors have appropriate security controls and documented ownership, and reduce dependency on vendors through internal knowledge.

Leadership Expectations

Operate as both a strategic technology leader and an effective technical decision-maker. Remain calm, organized, and decisive during outages, security events, and high-pressure situations. Create accountability without unnecessary bureaucracy. Communicate technical risks and recommendations clearly to technical and nontechnical audiences. Build productive partnerships across Engineering, Product, Data, Marketing, Security, and business operations. Make decisions based on reliability, risk, customer impact, cost, and measurable outcomes. Develop leaders who can independently manage their functions while maintaining consistent standards, and encourage teams to solve root causes rather than rely on repeated manual intervention.

Requirements

10+ years in cloud infrastructure, SRE, platform engineering, DevOps, IT operations, or related disciplines. 5+ years leading engineering or technology teams, including managers or technical leads. Demonstrated experience operating high-traffic, customer-facing platforms in AWS or a comparable cloud. Strong understanding of cloud architecture, networking, identity, security, observability, databases, APIs, storage, and distributed systems. Experience managing production availability, incident response, root-cause analysis, disaster recovery, and business continuity. Experience with infrastructure automation, configuration management, CI/CD, and infrastructure as code. Experience managing corporate IT, employee support, endpoint management, licensing, and access provisioning. Experience establishing operational metrics, SLOs, and support standards. Demonstrated ability to manage vendors, contracts, budgets, and cloud costs. Strong written and verbal communication skills, with the ability to explain technical issues and tradeoffs to executives., Experience with Cloudflare, DataDome, or comparable bot-management and web application protection platforms, and leading bot mitigation or traffic-quality programs. Experience with AWS services, containerized environments, infrastructure as code, centralized logging, and APM (e.g., AppDynamics). Experience supporting large-scale consumer subscription, public-record, data-as-a-service, advertising, or ecommerce platforms. Experience modernizing legacy infrastructure, transitioning systems from external vendors, and consolidating cloud accounts or IT services. Familiarity with privacy requirements and secure handling of sensitive consumer data. Experience building or maturing SRE, DevOps, or platform engineering functions.

Benefits & conditions

Posted Yesterday Remote Hiring Remotely in USA 170K-180K Annually Expert/Leader Remote Hiring Remotely in USA 170K-180K Annually Expert/Leader Leads cloud infrastructure, SRE, production operations, corporate IT, helpdesk, and bot mitigation. Owns AWS architecture, reliability, security, observability, incident response, disaster recovery, identity management, vendor relationships, budgets, and cloud costs. Builds operational standards, SLOs, automation, infrastructure roadmaps, and high-performing teams while improving customer-facing platform availability, traffic protection, employee technology support, and cybersecurity readiness. The summary above was generated by AI

PeopleFinders.com, the premier online service for consumers to locate, contact and verify people and businesses. Over the past couple of decades the Company has quietly become one of the largest owners of public records data in the country, distributing its products over a vast network of websites.

PeopleFinders Remote

Salary - $170k - $180k, 401k 401k match Medical/Dental/Vision/ Life Insurance, Remote or Hybrid United States 97K-131K Annually Senior level 97K-131K Annually Senior level Cloud * Fintech * Software * Business Intelligence * Consulting * Financial Services Leads end-to-end quality assurance for Workday implementations, upgrades, and enhancements. Defines testing strategies, plans, frameworks, and risk mitigation; oversees functional, integration, regression, and UAT testing across Workday modules. Manages defects, test data, environments, release validation, reporting, and stakeholder sign-off. Coordinates QA activities among business teams, developers, system integrators, vendors, and delivery partners while ensuring adherence to quality standards and governance. Top Skills: Azure DevopsErpJIRAWorkday Wipfli

Tax Manager

26 Minutes Ago Remote or Hybrid 106K-160K Annually Senior level 106K-160K Annually Senior level Cloud * Fintech * Software * Business Intelligence * Consulting * Financial Services Manage tax compliance engagements, supervise staff, and ensure timely communication with clients while maintaining technical tax expertise. Top Skills: AdobeCasewareExcelIdeaPowerPointProsystemsRiaWord, Wipfli, Remote or Hybrid Senior level Senior level Cloud * Fintech * Software * Business Intelligence * Consulting * Financial Services Manage complex tax returns for partnership clients, evaluate partnership agreements, mentor staff, and ensure exceptional client service. Requires strong leadership in tax consulting and compliance. Top Skills: Tax-Related Software

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.builtincolorado.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:40 min

Overcoming modern anti-bot mechanisms and network access restrictions

Vidas Bacevičius Vidas Bacevičius · World Congress 2025

1:05 min

Establishing access through laptop farms and unwitting facilitators

George Proorocu George Proorocu · World Congress 2026 Europe

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

2:18 min

Implementing anti-bot protocols with Privacy Access Control Tokens

Chris Heilmann Chris Heilmann +2 · LIVE

Videos

See all

Related articles

See all