> Markdown version of [/jobs/ext/2722612-infrastructure-engineer-kubernetes-specialist](https://www.wearedevelopers.com/jobs/ext/2722612-infrastructure-engineer-kubernetes-specialist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Engineer, Kubernetes Specialist - **Company:** Typesafe AI Inc. - **Location:** San Francisco, CA, United States - **Experience:** Experienced - **Salary:** $180,000.0 - $280,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Big Data, Data Infrastructure, Data Systems, Cursor (Graphical User Interface Elements), Software Debugging, Python (Programming Language), Next.js, Software Engineering, Data Streaming, TypeScript, Management of Software Versions, Tailwind, Large Language Models, Backend, Build Management, Kubernetes, Machine Learning Operations, Front End Software Development - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/infrastructure-engineer-kubernetes-specialist-typesafe-ai-8295157 ## About the Role * Have previously built big things * Have 5+ years of professional software engineering experience (3+ years working on backend, data infrastructure, or ML systems) * Have experience designing systems that acquire, transform, and manage large datasets efficiently * Have experience with large-scale data analysis and experimentation * Have experience working closely with researchers or ML engineers to translate data needs into scalable systems * Have experienced LLMs' capabilities and limitations from implementing them in the past ## Description As a data infrastructure engineer, you will build the internal data platform and tooling that powers TypeSafe's model training, evaluation, and experimentation. The systems you build will sit on the critical path of how we acquire data, evaluate model behavior, and improve our models through tight iteration loops. This data infrastructure is where much of the leverage in modern AI systems comes from! This role focuses on building the developer-facing systems that make working with large datasets and model outputs easy, safe, and scalable. Our tech stack is primarily Python. We also use TypeScript, Next.js, and Tailwind CSS for frontend, with Kubernetes for orchestration. We empower developers to use any tooling they find helpful for getting their job done, including Claude Code and Cursor. Responsibilities The role is wide ranging and you will wear many hats. Responsibilities include: * Design and build internal tools for managing datasets, model outputs, and evaluation results * Create reliable systems for dataset versioning, lineage, and reproducibility * Develop abstractions and APIs that allow research and product teams to interact with data without needing to understand underlying infrastructure * Build tooling that accelerates data acquisition, labeling, curation, and analysis * Create observability and debugging tools that make it easy to understand how data flows through training and evaluation systems * Collaborate closely with research, product, and infrastructure teams to ensure data systems support rapid experimentation while maintaining correctness and reliability We expect that you * Are responsible, ownership-inclined, and a team player - you believe there is no such thing as "other people's code" * Enjoy building tools and abstractions that improve developer productivity * Have experience designing data systems or internal developer platforms * Care deeply about correctness, reliability, and maintainability in systems that handle critical data * Collaborate well with others on technical and product design, advocating for what you need and adjusting to changing requirements * Are comfortable working in ambiguous problem spaces and building systems from first principles * Are mission aligned and excited to go all-in * Love being part of a team ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [GraphQL + Apollo + Next.js: A Lovely Trio](https://www.wearedevelopers.com/videos/311-graphql-apollo-next-js-a-lovely-trio) - [Developer Experience, Platform Engineering and AI powered Apps](https://www.wearedevelopers.com/videos/990-developer-experience-platform-engineering-and-ai-powered-apps) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) - [You are not an AI developer](https://www.wearedevelopers.com/videos/1148-you-are-not-an-ai-developer) ## Related Articles - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 159: AI Pipelines, 10x Faster TypeScript, How to Interview](https://www.wearedevelopers.com/magazine/563-dev-digest-159-ai-pipelines-10x-faster-typescript-how-to-interview) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)