> Markdown version of [/jobs/ext/628703-data-center-facility-telemetry-controls-engineer](https://www.wearedevelopers.com/jobs/ext/628703-data-center-facility-telemetry-controls-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Center Facility Telemetry & Controls Engineer - **Company:** Lambda Inc. - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $185,000.0 - $290,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Application Integration Architecture, Data Centers, Data Sharing, Data Center Infrastructure Management (CIM), Monitoring of Systems, Python (Programming Language), Modbus, Ansible, Prometheus, OPC Unified Architecture, Remote Infrastructure Management, Simple Network Management Protocols, Transmission Control Protocol (TCP), Grafana, Mttr, Influxdb, Bacnet, Api Design, Restful APIs, Terraform - **Published:** June 24, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=0de4abb25e2daf58 ## About the Role Do you have experience in Uninterruptible power supplies (UPS)?, * 7+ years of experience in data center infrastructure engineering, with at least 4 years focused on BMS, DCIM, or controls systems in a hyperscale, colocation, or AI/HPC environment. * Hands-on experience designing and integrating BMS for mission-critical facilities including UPS, PDU, CRAH/CRAC, chiller plant, cooling tower, and liquid cooling (CDU/in-row) systems. * Strong working knowledge of industrial control protocols: BACnet IP/MS-TP, Modbus TCP/RTU, SNMP, DNP3, and modern API-based integrations. * Demonstrated experience with DCIM platforms (Nlyte, Sunbird, Vertiv TRELLIS, or equivalent) including deployment, configuration, and ongoing administration. * Experience with real-time telemetry stacks (Prometheus, InfluxDB, Grafana, or similar) applied to infrastructure monitoring use cases. * Strong understanding of data center power and cooling systems, including PUE optimization, thermal management, and redundancy architectures (2N, N+1)., * Direct experience with direct liquid cooling (DLC) systems, CDU controls integration, and TCS loop management for high-density AI GPU deployments (100+ kW per rack). * Familiarity with OCP (Open Compute Project) hardware and telemetry standards. * Experience working with major colocation providers (Equinix, Digital Realty, CyrusOne, etc.) on BMS/EPMS integration and data sharing agreements. * Exposure to modular or edge data center deployments and associated controls considerations. * Background in scripting and automation (Python, Ansible, Terraform) applied to infrastructure management workflows. * Experience operating data centers at international scale, including Asia-Pacific or Southeast Asian markets. * Relevant certifications: CDCP, CDCE, ETA Data Center Specialist, or vendor-specific BMS/controls certifications. ## Description The Data Center Facility Telemetry & Controls Management Engineer is a critical technical role responsible for the design, deployment, integration, and ongoing operation of Building Management Systems (BMS), Data Center Infrastructure Management (DCIM) platforms, and facility telemetry pipelines across Lambda's growing data center portfolio. This engineer ensures that all facility systems - power, cooling, thermal, and environmental - are continuously monitored, alarmed, and controllable in real time, supporting the safe and efficient operation of high-density GPU deployments at rack densities of 136-380 kW per rack. What You'll Do BMS / Controls Architecture & Integration * Architect and manage BMS integration across colocation and Lambda-owned facilities, covering chillers, CRAHs, CDUs (Coolant Distribution Units), cooling towers, UPS systems, PDUs, and automatic transfer switches. * Define standards for BMS point lists, naming conventions, control sequences, and integration protocols (BACnet, Modbus, SNMP, OPC-UA, RESTful APIs). * Oversee commissioning and acceptance testing of new BMS deployments and CDU/TCS loop integrations for next-generation liquid-cooled GPU rack systems. * Collaborate with colocation partners (Equinix, Digital Realty, and others) to ensure telemetry data flows from provider BMS/EPMS into Lambda's monitoring stack. DCIM & Telemetry Platform Management * Own the DCIM platform strategy and roadmap - evaluating, selecting, and implementing tooling for asset management, capacity planning, environmental monitoring, and power chain visibility. * Develop and maintain real-time dashboards for PUE, thermal performance, stranded capacity, and cooling system efficiency across all Lambda sites. * Build and maintain telemetry pipelines ingesting data from BMS, PDUs, in-rack sensors, CDUs, and network devices into centralized monitoring and alerting platforms (e.g., Prometheus, Grafana, InfluxDB, or equivalent). * Define alarm thresholds and escalation workflows for critical facility events including high coolant temperatures, CDU inlet/outlet anomalies, leak detection, and power exceedances. Liquid Cooling Controls & High-Density Operations * Develop control strategies and setpoint frameworks for TCS (Thermal Control System) loops supporting direct liquid cooling at densities of 220-380 kW per rack. * Evaluate and qualify CDU vendors on controls integration capabilities, telemetry exposure, and remote management interfaces. * Define and enforce operational procedures for CDU commissioning, setpoint changes, loop pressure management, and fluid quality monitoring. * Support design and construction coordination for liquid cooling infrastructure in new data center buildouts, ensuring BMS and controls readiness at Day 1. Operational Reliability & Incident Response * Establish and maintain facility event management processes, including on-call response protocols for facility telemetry anomalies. * Lead root cause analysis for facility system failures and implement corrective actions to prevent recurrence. * Partner with the data center operations team to maintain and refine emergency response runbooks tied to BMS alerts and automated controls. * Drive continuous improvement in MTTR for facility-related events through better telemetry coverage and automated remediation. Vendor & Stakeholder Management * Manage BMS integrators, DCIM vendors, and control subcontractors - from RFP through design, installation, commissioning, and ongoing support. * Serve as the primary technical interface with colocation providers on all BMS/EPMS integration topics. * Collaborate with Lambda's infrastructure engineering, construction, and procurement teams to align controls requirements with facility buildout timelines. * Support due diligence and technical evaluation for new colocation sites and modular data center deployments from a telemetry and controls readiness perspective ## Related Videos - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Industrializing your Data Science capabilities](https://www.wearedevelopers.com/videos/178-industrializing-your-data-science-capabilities) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)