> Markdown version of [/jobs/ext/2215364-technical-architect](https://www.wearedevelopers.com/jobs/ext/2215364-technical-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Technical Architect - **Company:** Involvesthis - **Location:** Epsom, UK - **Salary:** £169,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon S3, Computer Vision, Continuous Integration, Systems Integration, UML, Management of Software Versions, Archimate, Togaf, SC Clearance, Low Latency, ONNX (Open Neural Network Exchange) Format, Machine Learning Operations, TensorRT - **Published:** August 25, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/technical-architect/44980105 ## About the Role overview one - you'll need to know, in practice, how model serving actually works at scale, where inference bottlenecks show up, and how the levers between accuracy, latency and cost actually trade off against each other. You should also be comfortable enough with computer vision (CNNs, vision transformers, detection and segmentation methods) to push back on a data science team's approach when needed. Day to day, you'd be:Taking ownership of the model lifecycle - pipelines, registry, versioning, CI/CD, drift detection, and retraining logicWorking out how to run models in constrained or edge settings: quantisation, pruning, distillation, ONNX/TensorRT, and picking the right accelerators for the jobSetting the evaluation approach - where precision and recall trade off, where thresholds sit, and how much false-positive friction is acceptable for the people using the systemDesigning so the model supports human decision-making rather than overrides itBuilding for sites with poor or intermittent connectivity, with training centralised and models pushed out from thereGetting the non-functional side right at scale - throughput, latency, availability, resilience, DR - across multiple sitesReviewing and challenging supplier architecture, including calling out vendor lock-in risk and build-vs-buy callsWorking to Secure by Design principles and NCSC guidance throughout, given the OFFICIAL-SENSITIVE classification What you'll bringA track record as a solution/technical architect on AI projects sitting inside bigger strategic programmesSolid AWS experience - S3, Lambda, EventBridge, SNS/SQSComfortable working across TOGAF, ArchiMate, C4, UML and BPMNHands-on with MLOps tooling and platforms - SageMaker, OpenVINO or equivalentExperience architecting data for large volumes of imagery - tiering, retention, lineage, provenanceFamiliar with the governance side of AI - DPIAs, model documentation, bias and fairness checksStrong stakeholder handling, and the ability to explain architecture to people who aren't architectsThis role requires active SC Clearance Bonus points forBackground in X-ray-based AI models or density/object identificationPublic sector delivery experience - GDS standards, Technology Code of Practice, spend controlsKnowledge of the DSIT AI Playbook or the Algorithmic Transparency Recording StandardFinOps experience specifically around GPU and inference costExperience integrating with scanner/hardware OEM systems and real-time image pipelines ## Related Videos - [Reference Architecture of AI in the Cloud](https://www.wearedevelopers.com/videos/1613-reference-architecture-of-ai-in-the-cloud) - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [ChatGPT and Java: A Match Made in Heaven or Hell?](https://www.wearedevelopers.com/videos/536-chatgpt-and-java-a-match-made-in-heaven-or-hell) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Beyond UML: Making Sense of AI-Generated Code through Visual Architecture](https://www.wearedevelopers.com/videos/2102-beyond-uml-making-sense-of-ai-generated-code-through-visual-architecture) ## Related Articles - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)