> Markdown version of [/jobs/ext/2258731-founding-engineer-agent-systems](https://www.wearedevelopers.com/jobs/ext/2258731-founding-engineer-agent-systems). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Founding Engineer, Agent Systems - **Company:** C&d Talent Advisory - **Location:** Hitchin, UK - **Experience:** Expert - **Salary:** £130,000.0 - £170,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Microsoft Azure, Cyber Security, Github, PostgreSQL, Node.Js, Regression Testing, Software Deployment, TypeScript, Management of Software Versions, Enterprise Software Applications, Tailwind, ReactJS, Large Language Models, Multi-Agent Systems, IT Architecture, Backend, Terraform, Docker - **Published:** August 26, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5854877274 ## About the Role * Strong professional backend engineering experience * 1-2+ years of experience shipping production LLM-powered features * Strong TypeScript or Node.js experience * Experience building customer-facing production AI systems * Hands-on experience with agent frameworks * Experience with tool calling and multi-step orchestration * Production experience building AI evaluations * Experience with dataset curation and regression testing * Understanding of LLM-as-judge approaches and failure modes * Strong systems-thinking ability * Experience with asynchronous systems and queues * Understanding of idempotency and reliability patterns * Strong product-engineering mindset * Ability to own features from customer problem through production deployment * High level of autonomy and accountability * Based in London or able to commute regularly for an on-site-first role Particularly valuable experience * Anthropic APIs * OpenAI APIs * Open-weight models * Production LLM workloads at scale * Agent orchestration frameworks * AI evaluation infrastructure * Model fallback architecture * Prompt versioning * Prompt-injection defense * Agent security * Sandboxed tool execution * Production RAG systems * High-stakes AI applications * Compliance * Governance, Risk and Compliance (GRC) * Audit technology * Cybersecurity * Financial services * Regulated enterprise software Strong profile signals * Former founder or technical founder * Former CTO returning to hands-on engineering * Founding or early-stage engineering experience * Experience at a high-growth product company * Experience at a leading technology or AI organization * Significant independent AI projects or side projects * Strong evidence of product ownership * History of shipping customer-facing systems end-to-end * AI-native development workflow * Heavy use of coding agents such as Claude Code or equivalent * Ability to move between product, backend, AI systems, and reliability work * Strong technical judgment rather than purely research-oriented expertise This role may not be the right fit if you * Have exclusively front-end or front-end-first engineering experience * Have a purely AI research background without production software engineering ownership * Have worked primarily in Developer Relations or Developer Experience * Have only built internal AI tools rather than customer-facing production systems * Lack meaningful backend engineering experience * Have AI experience limited to prototypes or demos * Prefer research over shipping production products * Lack experience taking ownership of production reliability * Prefer narrowly scoped engineering responsibilities * Are uncomfortable making judgment calls around AI quality and production readiness * Are not comfortable working primarily on-site in London Technical environment ## Description * Develop secure sandboxing for agent execution * Implement prompt-injection defenses and other agent-security controls * Build production evaluation frameworks for fuzzy, high-stakes outputs * Curate evaluation datasets and regression suites * Design and improve LLM-as-judge systems * Evaluate model performance across upgrades and provider changes * Build retries, fallbacks, and circuit-breaker mechanisms * Own prompt versioning and AI-quality infrastructure * Design systems around async processing, queues, and idempotency * Define the internal standard for production AI quality * Own AI reliability and make ship/no-ship decisions when necessary, * Node.js * React * Tailwind * Express * Azure * Postgres * Terraform * GitHub Actions * Docker * Anthropic-first AI stack * Claude Code What you'll get * Founding-level ownership of the agent platform * Significant influence over AI architecture and engineering standards * Opportunity to build production-grade agent systems from an early stage * Direct impact on customer-facing AI quality and reliability * Exposure to frontier LLM APIs in demanding production environments * Work on technically difficult evaluation and reliability problems * Close collaboration with company leadership * Top-decile London market compensation * Meaningful EMI-eligible equity * Significant AI tooling and API budgets * Daily team lunch * Specialty coffee * King's Cross office with roof terrace and on-site facilities * Fast-moving engineering environment with high autonomy ## Related Videos - [Watch Tests Go Brrrr! : Getting Started with Cypress in ReactJS](https://www.wearedevelopers.com/videos/282-watch-tests-go-brrrr-getting-started-with-cypress-in-reactjs) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Agentic employees in world's most downloaded FinTech app](https://www.wearedevelopers.com/videos/100123-agentic-employees-in-world-s-most-downloaded-fintech-app) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Headless by Design: Building Enterprise Systems That Agents Can Actually Use](https://www.wearedevelopers.com/videos/100092-headless-by-design-building-enterprise-systems-that-agents-can-actually-use) ## Related Articles - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Never delegate the understanding](https://www.wearedevelopers.com/magazine/749-never-delegate-the-understanding) - [What is Agentic Programming and Why Should Developers Care?](https://www.wearedevelopers.com/magazine/625-what-is-agentic-programming-and-why-should-developers-care) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)