AI Infrastructure Data Center Deployment Lead

Lambda Inc.
Kansas City, MO, United States
12 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$109,000.0 - $145,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence JIRA Big Data Computer Engineering Data Centers Network Configuration and Change Management Software Engineering Systems Architecture AI Infrastructure High Performance Computing Deployment Automation Machine Learning Operations

Job description

Lambda, Inc. is seeking a highly skilled and experienced data center technician to join our Infrastructure deployment team. As a member of AI Infrastructure Deployments you will be the on-site face of Lambda to vendors and contractors for data center deployments. You will be responsible for completing deployments and improvements of Lambda’s integrated data center solutions. You will collaborate with your teammates and cross-functional teams to pilot new technologies, physically deploy large AI/ML clusters, setup network configurations, and build out Lambda data center locations. Additionally, you will interface with Hardware Engineering and Supply Chain teams to comply with operational standards and requirements to improve deployment quality and speed. Your focus will be on becoming skilled at both datacenter builds for high performance computing environments as well as the deployment of AI/ML systems.

Data Center Deployments:

  • Be a primary deployment member for major cluster / data center deployments as well as be able to respond to smaller on-premise deployments, if needed.
  • Develop your skillset for new data center deployments, deployment planning, and technical AI infrastructure and cluster configurations.
  • Be responsive to project management, Supply Chain, HPC logical configuration, and Data Center Operations teams during deployment planning and execution.
  • Collaborate with cross-functional teams to address deployment-related challenges and optimize regional operations.
  • Interface with vendors to guide the build out of Lambda data center to meet exacting standards.
  • Understand power/cooling requirements as well as cabling needs required within data center space to support high performance infrastructures to guide data center build outs.
  • Participate in technical discussions and provide input on data center integration and deployment strategies.
  • Work closely with cross-functional teams, including Hardware Engineering, Software Engineering, Supply Chain, Customer Experience and Sales, to align data center solutions with business goals.

AI / ML Infrastructure:

  • Be conversant on infrastructure components (compute, storage, networking) used in deployments. Understanding component configuration and integration to guide system deployments, design, and operations.
  • Collaborate with stakeholders to understand customer requirements and translate them into actionable plans.
  • Interface with Hardware Engineering and Architecture teams to develop operational standards and requirements for infrastructure deployments.

Operational Standards and Requirements:

  • Follow best practices and guidelines for efficient and consistent data center operations.
  • Ensure compliance with industry standards and regulations in datacenter safety.

Continuous Improvement and Innovation:

  • Identify opportunities for process improvement and operational efficiency in data center integration and deployment.
  • Stay updated with the latest industry trends, technologies, and best practices.
  • Drive innovation and propose new solutions to enhance data center operations.

Requirements

  • Experience in data center operations and integration, preferably in a leadership role.
  • Experience with large scale data center and infrastructure deployments.
  • Understanding of data center design, deployment, and architecture.
  • Understanding of deployment planning.
  • Knowledge of industry standards and regulations related to data center operations.
  • Problem-solving and analytical skills, with the ability to address data center challenges.
  • Communication and interpersonal skills to effectively collaborate with cross-functional teams and stakeholders.
  • Provide guidance to numerous teams involved in systems deployments.

Nice to Have

  • Familiarity with hardware engineering, supply chain, and inventory management processes. (NetBox, Jira, etc.)
  • Knowledge of high performance computing technologies and their integration with data center operations.
  • Expertise with DGX & HGX based systems architecture.
  • Professional certifications related to data center operations and integration.

Benefits & conditions

Pulled from the full job description Health insurance 401(k) matching Paid time off Vision insurance Dental insurance, This is a salaried non-exempt role, eligible for overtime. The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, with 500+ employees, and growing fast
  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
  • Our values are publicly available: https://lambda.ai/careers
  • We offer generous cash & equity compensation
  • Health, dental, and vision coverage for you and your dependents
  • Wellness and commuter stipends for select roles
  • 401k Plan with 2% company match (USA employees)
  • Flexible paid time off plan that we all actually use

About the company

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda’s mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you’d like to build the world’s best AI cloud, join us.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

2:03 min

Solving complex engineering challenges in artificial intelligence deployment

Nico Axtmann · WWC 2022

Videos

See all

Related articles

See all