> Markdown version of [/jobs/ext/3048277-principal-architect-distributed-systems-event-streaming-director-i-product-architect](https://www.wearedevelopers.com/jobs/ext/3048277-principal-architect-distributed-systems-event-streaming-director-i-product-architect). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Architect - Distributed Systems & Event Streaming (Director I - Product Architect) - **Company:** Ust View All Jobs - **Location:** Nottingham, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Audit Trail, Code Review, Continuous Integration, Command-Query Responsibility Segregation (Software Development), Serialization, DevOps, Distributed Systems, Fault Tolerance, Reliability Engineering, Prometheus, Software Engineering, Data Streaming, Cloud Platform System, Performance Testing, Real Time Systems, Grafana, Concurrency, Apache Spark, Rate Limiting, Event Driven Architecture, Kubernetes, Data Lineage, Low Latency, Apache Flink, Production Code, Apache Kafka, Spark Streaming, Asynchronous Programming, Event Sourcing, Stream Processing, Dynatrace - **Published:** September 24, 2026 - **Apply:** https://www.careerjet.co.uk/job/gb926441a25c7f912f6693b1175d08e644/eaa ## About the Role * 10+ years of experience in software engineering, distributed systems, platform engineering, or architecture (12-18 years preferred). * Proven ownership of large-scale, production distributed systems. * Deep hands-on expertise with Kafka, Redpanda, or comparable event-streaming platforms. * Strong experience designing and operating event-driven architectures in enterprise environments. * Expert-level Java engineering skills. * Deep understanding of: * Distributed systems design * Consistency models * Ordering and idempotency * Concurrency and backpressure * Throughput and latency optimisation * Replication, partitioning, and failure recovery * Experience building resilient systems using retries, circuit breakers, bulkheads, dead-letter queues, replay, rate limiting, and graceful degradation patterns. * Strong debugging skills across code, messaging, infrastructure, observability, and production operations. * Excellent communication and mentoring capability across globally distributed teams. Preferred * Redpanda architecture and operations. * Event sourcing and CQRS in production environments. * Stream processing with Kafka Streams, Flink, Spark Streaming, ksqlDB, or similar technologies. * Financial services, banking, lending, payments, or regulated industry experience. * Kubernetes and cloud-native platforms. * Observability stacks including Prometheus, Grafana, OpenTelemetry, and distributed tracing. * CI/CD, contract testing, performance testing, chaos engineering, and production-readiness practices. * Event governance, auditability, data lineage, and security for event-driven platforms., Data Governance, Design Patterns, Distributed Systems, Event-Driven Architecture, ksqlDB, Apache Spark, Apache Kafka, Apache Flink, Observability, Redpanda, Streaming ## Description UST FinX is seeking a Principal Architect to lead the technical direction of event-driven architecture and distributed systems across our platform. This is a senior, hands-on engineering leadership role where influence comes from building resilient systems, writing production-grade code, creating reusable frameworks, solving complex technical challenges, and raising engineering standards across the organisation. You will own the architecture and evolution of our event platform, including Kafka/Redpanda, event contracts, schema management, event sourcing, CQRS, distributed consistency, high-throughput processing, resilience, observability, and operational excellence. This is not an advisory or documentation-focused architecture role. We are looking for a practitioner who stays close to implementation and is passionate about building scalable, reliable systems in production. What You'll Do * Define and evolve the event-driven architecture strategy for UST FinX. * Lead architecture across distributed systems, event streaming, integration, orchestration, and real-time processing services. * Build reusable frameworks, libraries, and reference implementations for producers, consumers, retries, replay, resilience, and observability. * Establish standards for event contracts, schema evolution, compatibility, governance, and lifecycle management. * Guide teams on Kafka/Redpanda architecture, including topic design, partitioning, consumer groups, throughput, replay, ordering, and failure handling. * Apply event sourcing, CQRS, asynchronous communication, and eventual consistency where they provide business value. * Remain hands-on with Java, distributed systems, messaging infrastructure, and platform engineering. * Troubleshoot complex production issues involving latency, ordering, consumer lag, duplication, retries, serialization, and distributed state. * Mentor engineers through architecture reviews, design discussions, code reviews, and incident analysis. * Partner closely with Engineering, Product, Platform, Data, DevOps, QA, and Delivery teams to ensure systems are secure, scalable, observable, and production-ready. Reporting directly to the CTO, you will act as one of UST FinX's most senior technical leaders and a trusted force multiplier across engineering teams. You will influence multiple product and platform pods, establish reusable engineering standards, challenge weak designs, and ensure architectural decisions are reflected in real systems, code, and operational outcomes. Your authority will come from technical depth, engineering judgement, and execution rather than organisational hierarchy. What We're Looking For ## Related Videos - [Kafka Streams Microservices](https://www.wearedevelopers.com/videos/168-kafka-streams-microservices) - [The Power of Purpose: Unlocking Potential and Innovation](https://www.wearedevelopers.com/videos/1110-the-power-of-purpose-unlocking-potential-and-innovation) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [It's Not Vibe Coding If You Know What You're Doing](https://www.wearedevelopers.com/videos/100119-it-s-not-vibe-coding-if-you-know-what-you-re-doing) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Why Event-Driven Architecture Isn’t About Speed (and When You Actually Need It)](https://www.wearedevelopers.com/magazine/745-why-event-driven-architecture-isn-t-about-speed-and-when-you-actually-need-it) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)