> Markdown version of [/jobs/ext/2729076-ai-engineer](https://www.wearedevelopers.com/jobs/ext/2729076-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI engineer - **Company:** Ai - **Location:** Berlin, Germany - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Cloud Computing, Customer Data Management, Statistical Hypothesis Testing, Python (Programming Language), Key Management, PostgreSQL, Regression Testing, Software Engineering, TypeScript, AI Infrastructure, Large Language Models, Model Validation, Caching, Backend, Kubernetes, Production Code, Api Design, Docker, Programming Languages - **Published:** September 5, 2026 - **Apply:** https://startup.jobs/founding-ai-engineer-embedded-ai-efficiency-layer-berlin-atlanticvc-9774866 ## About the Role You're an experienced AI engineer who combines strong software engineering with original thinking about how production AI systems can become more efficient. You can reason from first principles about where cost and quality are lost, test new approaches rigorously and turn the strongest ideas into production-grade code. * You have substantial experience in software engineering, backend systems and API development. * You are highly proficient in python or TypeScript or at least some coding languages and comfortable working across them. * You understand the foundations of modern LLM systems, including tokenisation, context management, inference behaviour, model evaluation and the trade-offs between cost, latency and quality. * You have hands-on experience building with LLM APIs and understand the behaviour, cost and reliability challenges of production AI systems. * You have practical experience designing experiments and evaluations that measure both model quality and system performance. * You have worked with cloud infrastructure, PostgreSQL and Docker. Kubernetes experience is a strong advantage. * Experience with vLLM, SGLang, LiteLLM or similar AI infrastructure is highly relevant. * You understand distributed-system performance, including latency, throughput, caching, rate limits and failure handling. * You can independently research difficult problems, test hypotheses and translate technical ideas into working systems. * You write clear, reliable code and are comfortable owning systems through deployment and production. * Fluent English is required. German is an advantage. We do not expect you to have worked on every system or technology listed above. We do expect deep experience in some of these areas, strong software engineering fundamentals and the ability to develop genuine technical depth in the rest. ## Description You'll join as our first AI engineer and work directly with the founders to build the core product from the ground up. You'll own major areas of product development, help define our architecture and technical direction, and turn complex research and engineering problems into reliable production systems. This role requires someone who can think deeply about AI systems and write the code to build them. You'll work across AI infrastructure, backend systems, evaluation and enterprise deployments, with substantial ownership over the technical decisions behind the product. What You'll Do * Research, invent and productionise systems for AI cost efficiency, context optimisation and token reduction. * Build across context selection, compression, caching, model routing, tool-call reduction, retry control and output budgets. * Develop APIs, an OpenAI-compatible gateway and integrations with model providers and enterprise AI applications. * Create reproducible cost-quality evaluations, regression tests and quality gates. * Own production reliability across monitoring, deployments, incidents, capacity, availability and latency. * Build secure enterprise infrastructure with tenant isolation, access controls, secrets management and protected customer data. * Optimise distributed-system performance across throughput, caching, rate limits, resource usage and failure handling. * Translate early customer requirements into reliable product capabilities. * Take on the broad day-to-day engineering work required to move an early product forward. ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [Three years of putting LLMs into Software - Lessons learned](https://www.wearedevelopers.com/videos/1508-three-years-of-putting-llms-into-software-lessons-learned) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [You are not an AI developer](https://www.wearedevelopers.com/videos/1148-you-are-not-an-ai-developer) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)