Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG, INC.
New York, NY, United States
about 1 month ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$190,000.0 - $290,000.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Data Infrastructure Python (Programming Language) Machine Learning Data Streaming Data Storage Technologies Apache Kafka Stream Processing

Job description

As a Senior ML Engineer for AWS and Real-Time Inference, you’ll own the fast path: ingesting live trading data and scoring it in near real time. It’s a systems-heavy role focused on streaming, low-latency inference, and the retraining cadence that keeps models current and most directly determines whether the systems keep up with live markets., * Streaming and storage pipelines that feed both model training and low-latency inference.

  • The online inference path and its latency. A real-time detector microservice is built and unit-tested but not yet deployed - it needs to be connected to a run-time model and hold latency under live load. One known, non-trivial problem lives here: batch scoring ranks across a whole population, but single-account (or single-wallet) real-time scoring has no population to rank against, so it must threshold on calibrated raw scores.
  • The model retraining cadence as data and labels accumulate, including drift-triggered retraining.
  • Productionizing new features and detectors on the fast path, in partnership with data science.

Requirements

  • Strong data / ML engineering experience with streaming systems (e.g., Kafka / Kinesis / MSK) and modern data storage formats
  • Experience building low-latency, high-throughput inference services
  • Proficiency in a systems language (e.g., Go) alongside Python
  • Production AWS experience
  • Familiarity with financial market data or trading protocols a plus - the US feed is a FIX 5.0 SP2 drop-copy session with real-world quirks (nanosecond timestamps, repeating groups, dedup semantics)
  • Familiarity with chain-data infrastructure (node providers, subgraphs, event indexing) is a plus

Benefits & conditions

The base pay for this position is $190,000-290,000. A bonus will be provided as part of the compensation package, in addition to a full range of medical, financial, and/or other benefits.

About the company

At TWG AI, we drive innovation and business transformation across a range of industries-including financial services, insurance, technology, media, and sports-by leveraging data and AI as core assets. Our AI-first, cloud-native approach delivers real-time intelligence and interactive business applications, empowering informed decision-making for both customers and employees.

We prioritize responsible data and AI practices, ensuring ethical standards and regulatory compliance. Our decentralized structure enables each business unit to operate autonomously, supported by a central AI Solutions Group, while strategic partnerships with leading data and AI vendors fuel game-changing efforts in marketing, operations, and product development.

You will collaborate with management to advance our data and analytics transformation, enhance productivity, and enable agile, data-driven decisions. By leveraging relationships with top tech startups and universities, you will help create competitive advantages and drive enterprise innovation.

At TWG, your contributions will support our goal of sustained growth and superior returns, as we deliver rare value and impact across our businesses.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:09 min

Evaluating mature stream processing frameworks for production systems

Soroosh Khodami Soroosh Khodami · World Congress 2024

1:37 min

Introduction to Apache Kafka benchmarking and performance analysis

Kirill Kulikov · LIVE

1:50 min

Lowering pipeline latency with data streaming

Nathaniel Okenwa Nathaniel Okenwa · World Congress 2024

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

2:13 min

Modernizing legacy applications for real-time streaming data consumption

Farooq Sheikh Farooq Sheikh +3 · World Congress 2025

50 sec

Identifying real-world applications for stream processing technology usage

Soroosh Khodami Soroosh Khodami · World Congress 2024

Videos

See all

Related articles

See all