Founding Cloud Infrastructure Engineer

AI AGENTS LLC
San Mateo, CA, United States
3 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Distributed Systems Fault Tolerance Python (Programming Language) Apache Spark Kubernetes Infrastructure Automation Frameworks Dask Apache Kafka
+1 more
Terraform

Job description

  • Architect & Scale Infra: Design and deploy secure, fault-tolerant cloud infrastructure across AWS/Azure/GCP, using Kubernetes, Terraform, and modern IaC tools.
  • Enable Distributed AI Workloads: Build and optimize compute systems for distributed frameworks like Ray and Kafka, ensuring scalability and reliability.
  • GPU Orchestration: Manage GPU resources and scheduling for ML inference, retrieval, and training workloads.
  • Drive Observability & Reliability: Implement monitoring, security, and best practices for high-availability ML/AI systems.
  • Collaborate Across Teams: Work hand-in-hand with ML researchers and data scientists to ensure infra accelerates, not hinders, development.

Requirements

  • We’re open to multiple levels of seniority (3-10+ years of engineering experience). What matters most is founder mindset + strong infra chops.
  • Hands-on with Kubernetes + Terraform (real-world deployments, not just exposure).
  • Experience with distributed systems (Ray, Dask, Spark, or similar).
  • Exposure to Kafka or equivalent MQ systems.
  • Strong programming/scripting in Python, Go, or similar.
  • Track record of building or operating production-grade infra. Nice to Haves

  • GPU scheduling/optimization experience.
  • Prior startup/early-stage build-from-scratch experience.
  • Experience supporting ML/AI workloads in production.
  • Founder Traits
  • Excited to own infra end-to-end as the first dedicated hire.
  • Comfort with ambiguity; thrives in fast-moving, collaborative teams.
  • Curiosity + grit: hungry to learn where gaps exist.

Benefits & conditions

  • Competitive salary, benefits, and meaningful equity.
  • Work alongside engineers and researchers from LinkedIn, Visa, Meta, and Branch.
  • Onsite culture in San Mateo, designed for deep collaboration and high-velocity building.
  • We offer competitive compensation, equity, full benefits (Medical, Dental, Vision, 401k)
  • We sponsor H-1B visas and assist with immigration processes We value builders over résumés. If this role excites you but you don’t check every box, we still want to hear from you. zaimler is an equal opportunity employer.

About the company

About zaimler AI agents can’t reason over data they don’t understand. Enterprise data today is fragmented across dozens of systems with no shared context, meaning, or structure, and that’s why most enterprise AI is failing. The shift from copilots to autonomous agents is creating an entirely new infrastructure layer, and we’re building it. zaimler is the context infrastructure for the agentic era: a platform that automatically discovers domain knowledge, maps relationships, and gives AI agents the semantic understanding to operate with precision at scale. Imagine knowledge graphs that support real-time inference, built for systems that need to reason, not just retrieve. zaimler was founded by Biswajit Das (ex-VP Engineering, Truera), a Data Infra veteran and former Chief Architect at Visa, and Sofus Macskassy (ex-Director of Engineering, LinkedIn), who built one of the largest knowledge graphs in production in the industry at LinkedIn. We’re growing and deploying with major enterprises across insurance, travel, and technology. If you want to build infrastructure that the next decade of enterprise AI runs on, we’d love to talk. What You’ll Do As our Founding Cloud Infrastructure Engineer, you’ll design, build, and own the cloud infrastructure that powers zaimler’s semantic platform. This isn’t a maintenance role. It’s a foundational build-from-scratch position, shaping choices we’ll live with for years.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

2:19 min

Scaling performance across multiple GPUs using specialized frameworks

Paul Graham Paul Graham · World Congress 2025

1:49 min

Accelerating initial deployments using AI generation agents

Marcel Scherenberg Marcel Scherenberg · World Congress 2025

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all