> Markdown version of [/jobs/ext/2709028-data-center-operations-engineer](https://www.wearedevelopers.com/jobs/ext/2709028-data-center-operations-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Center Operations Engineer - **Company:** COLOVORE LLC - **Location:** Chicago, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, Computerized Maintenance Management Systems, Data Centers, Monitoring of Systems, Issue Tracking Systems, High Performance Computing, Information Technology, Hardware Infrastructure - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-data-center-operations-engineer-colovore-9007256 ## About the Role Technical Mastery of Mission-Critical Infrastructure * Deep understanding of electrical, mechanical, and IT infrastructure systems (power distribution, cooling, cabling, server hardware) * Ability to diagnose, troubleshoot, and resolve complex infrastructure issues under time-sensitive conditions * Skilled in operating and interpreting data from BMS, CMMS, ticketing, and monitoring tools Operational Excellence & Preventative Maintenance * Strong command of SOPs, MOPs, and EOPs with a disciplined approach to execution * Expertise in designing, managing, and improving preventative maintenance programs * Detail-oriented mindset for inspections, documentation, compliance, and audit readiness Incident Leadership & Decision-Making * Ability to remain calm, clear-thinking, and decisive during outages or operational escalations * Strong situational awareness and risk assessment capabilities to drive safe, effective resolutions * Experience coordinating stakeholders and communicating clearly during high-pressure events Systems Optimization & Continuous Improvement * Analytical approach to tuning and optimizing power, cooling, and monitoring systems for high-density environments * Ability to identify inefficiencies, propose improvements, and drive implementation across teams * Comfortable working with data to inform operational decisions and capacity planning Customer Service & Communication * Professional, customer-oriented approach when supporting remote hands, escalations, or service recovery * Clear written and verbal communication skills, able to translate technical updates for both technical and non-technical audiences * Skilled at balancing customer needs with site-level operational priorities Leadership & Team Development * Experience training and mentoring junior technicians, with an emphasis on safety, accuracy, and professional growth * Ability to model operational discipline, set expectations, and ensure adherence to processes * Strong collaborator who works well across facilities, engineering, and customer-facing teams, * 5+ years of data center or mission-critical facility operations experience, with a strong focus on hands-on field work. * Proven expertise with power, cooling, and IT infrastructure systems in high-density environments. * Hands-on experience with BMS, CMMS, ticketing, and monitoring tools. * Demonstrated ability to lead incident resolution and high-pressure operational decisions. * Prior experience mentoring or leading junior technicians preferred. ## Description * Lead day-to-day operational support, including equipment installation, cabling, and facility rounds. * Oversee preventative maintenance cycles and ensure compliance with operational standards and audit requirements. * Manage and resolve escalations during incidents; act as decision-maker in high-pressure situations. * Operate and optimize Building Management Systems (BMS), Computerized Maintenance Management Systems (CMMS), and ticketing systems. * Configure, monitor, and tune infrastructure systems to improve reliability and efficiency. * Mentor and train junior operations staff; ensure proper documentation and SOP adherence. * Interface with strategic customers, providing remote hands support, escalation management, and service recovery. * Partner with facilities, engineering, and customer success teams to align on operational needs and improvements. * Maintain accurate records, logs, and reports for compliance, audits, and performance reviews. * Support capacity planning and operational readiness for new deployments., * Zero downtime achieved through proactive monitoring, preventative maintenance, and effective incident management. * Smooth customer experience with quick, accurate, and professional handling of remote hands requests and escalations. * High-performing systems with optimized power, cooling, and monitoring configurations that support dense AI/HPC workloads. * Well-trained team of operations staff that follows SOPs, executes work efficiently, and grows under mentorship. * Accurate, compliant documentation that meets audit requirements and provides transparency for leadership and customers. * Continuous improvements implemented in workflows, safety, and operational efficiency to scale with business growth. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [The Sustainability Race: AI's Promises, Pitfalls and Potential](https://www.wearedevelopers.com/videos/100155-the-sustainability-race-ai-s-promises-pitfalls-and-potential) - [APItoolkit: Using Merkle Trees and LLMs to Detect the UnDetectable in Software Monitoring](https://www.wearedevelopers.com/videos/1639-apitoolkit-using-merkle-trees-and-llms-to-detect-the-undetectable-in-software-monitoring) - [What makes Cybersecurity different for critical infrastructure?](https://www.wearedevelopers.com/videos/571-what-makes-cybersecurity-different-for-critical-infrastructure) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [System Resilience: Surviving the Software Storm](https://www.wearedevelopers.com/videos/874-system-resilience-surviving-the-software-storm) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [A Guide to Green Tech and Green IT Careers](https://www.wearedevelopers.com/magazine/374-a-guide-to-green-tech-and-green-it-careers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Best Paying Jobs in Technology](https://www.wearedevelopers.com/magazine/256-best-paying-jobs-in-technology) - [Data Science & more: The Lopez dilemma](https://www.wearedevelopers.com/magazine/10-data-science-more-the-lopez-dilemma) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)