Senior Platform Engineer - Real-Time Data & ML

Hanson Regan Ltd
Berlin, Germany
1 day ago
Apply on www.careerboard.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
€120,000.0
Working hours
Regular working hours

Tech stack

Cloud Computing Data Infrastructure Database Applications Distributed Systems Fault Tolerance Machine Learning Software Engineering Data Streaming Real Time Systems System Availability Event Driven Architecture Kubernetes
+7 more
Low Latency Apache Flink Real Time Data Apache Kafka Data Management Machine Learning Operations Vertica

Job description

We’re seeking an experienced Platform Engineer to help develop the infrastructure behind sophisticated, Real Time data and machine learning workloads.

You’ll be responsible for solving challenging engineering problems across distributed systems, streaming data and cloud infrastructure. The role will suit someone who enjoys working at scale, takes ownership of complex technical challenges and can build platforms that are both highly performant and dependable.

The Role:

  • Take end-to-end ownership of the infrastructure supporting Real Time machine learning and data-driven applications.
  • Develop highly scalable systems capable of handling significant volumes of continuously generated data while maintaining consistently fast response times.
  • Create platform components, tooling and engineering standards that make it easier for development teams to deliver and run reliable services.
  • Build and maintain event-driven architectures and Real Time data pipelines designed for demanding, latency-sensitive workloads.
  • Identify opportunities to improve system performance, resilience, scalability and operational efficiency.
  • Collaborate with engineers and specialists across machine learning, data, Back End and product to understand technical challenges and deliver effective solutions.
  • Contribute to architectural decisions and help shape the future direction of the platform.
  • Act as a technical role model within the team, supporting colleagues through mentoring, knowledge sharing and constructive engineering practices.

Requirements

  • Professional experience in platform, infrastructure, systems or software engineering, ideally within complex production environments.
  • Strong practical knowledge of Kubernetes and experience running services reliably in containerised environments.
  • A good grasp of distributed computing principles, networking, system performance, fault tolerance and designing for high availability.
  • Experience working with systems where latency and throughput are important considerations.
  • The ability to assess different technical approaches and make sensible decisions around scalability, reliability, performance and infrastructure costs.
  • Confidence operating across different technical areas, including cloud infrastructure, Back End services, data platforms and machine learning systems.
  • A pragmatic approach to engineering, with the ability to take loosely defined problems and turn them into clear, maintainable solutions.
  • Strong communication skills and the ability to work effectively with both technical and non-technical stakeholders.
  • A high degree of autonomy, curiosity and ownership, alongside a willingness to support the development of other engineers.

Desired Skills:

It would be advantageous if you have worked with technologies such as Kafka, Flink or ClickHouse, or similar tools used for event streaming, Real Time processing and large-scale analytical workloads.

Experience building platforms for low-latency applications, Real Time ML or high-volume data environments would also be highly relevant.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerboard.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

3:37 min

Accessing API documentation and testing remote driving latency

Alexandru Ciinaru Alexandru Ciinaru +3 · World Congress 2025

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all