AI ML - Senior Associate - Machine Learning Engineer

Reddit Inc.
London, UK
1 day ago
Apply on nlppeople.com
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Airflow Batch Processing BigQuery Data Infrastructure Programming Tools Distributed Data Store Machine Learning Video Encoding Feature Engineering Apache Spark Pyspark
+8 more
Kubernetes Code Testing Apache Flink Apache Kafka Machine Learning Operations Software Coding Stream Processing Data Pipelines

Job description

This is not a pure ML modeling role. The ideal candidate is excited about building reliable infrastructure, data pipelines, and developer-facing tools that make ML engineers more productive.

What You’ll Do

Design and build data infrastructure that supports large-scale feature and training set computation, transformation, and storage. Develop frameworks for batch and real-time features with a focus on reliability, scalability, and ease of use. Build platform capabilities for feature governance, including lineage tracking, validation, drift detection, anomaly monitoring, reproducibility, and versioning Partner with ML engineers to ensure smooth integration of feature engineering workflows into ML production systems. Build systems that support agentic ML workflows, including automated feature discovery, feature quality evaluation and feature lifecycle management Contribute to operational excellence through observability, performance tuning, reliability engineering, and cost optimization initiatives.

Requirements

We are looking for an engineer with experience in building high-scale data infrastructure and exposure to ML platforms to help evolve and scale our feature management systems., 3+ years in data infrastructure/platform engineering or ML infrastructure platforms. Hands-on experience building production services, data pipelines, APIs, workflow systems, or developer tools. Experience with at least one distributed data or compute system such as Spark, PySpark, Flink, Kafka, Ray, Airflow, Kubernetes, BigQuery, or similar technologies. Familiarity with ML data workflows such as feature generation, training dataset creation, batch processing, real-time data processing, model training, experimentation, or online serving. Strong coding skills and ability to write clean, maintainable, well-tested code. Experience building intelligent automation or agentic workflows for ML systems is a strong plus Experience with ML infrastructure and MLOps workflows spanning feature engineering, training pipelines, experimentation, model deployment, and online serving is a plus, Senior (5+ years of experience)

Tagged as: Industry, Machine Learning, NLP, United Kingdom

Benefits & conditions

Global Benefit programs that fit your lifestyle, from workspace to professional development to caregiving support Family Planning Support Gender-Affirming Care Mental Health & Coaching Benefits Group Personal Pension Scheme with Employer match Private Medical and Dental Scheme Income Replacement Programs Bike to Work scheme Flexible Vacation & Paid Volunteer Time Off Generous Paid Parental Leave

About the company

Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 126 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nlppeople.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:30 min

Leveraging BigQuery ML for scalable SQL-based segmentation experiments

Julian Joseph · LIVE

7:10 min

Exploring pathways into the machine learning engineering field

Jose Luis Latorre Millas · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all