Senior Machine Learning Engineer

Caterpillar
Mossville, IL, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$112,710.0 - $183,140.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Data Analysis Big Data Databases Data Infrastructure Extract Transform Load (ETL) Data Stores Data Visualization Relational Databases Programming Tools Linux System Administration Machine Learning
+8 more
NoSQL Raw Data Software Engineering Cloud Platform System Database Optimization Information Technology Data Management Data Pipelines

Job description

As a Machine Learning Infrastructure Engineer, you will be responsible for developing and deploying a robust full-stack pipeline that can support various perception/planning/prediction projects. In this role, you will build up various tools that will optimize efficiency when working with large data sets. This includes developing Data Stores, implementing Data Visualization utilities, and deriving quantifiable metrics to support the Machine Learning pipeline. Your actions will directly contribute to the development of safer, more efficient machines for Caterpillar’s customers. Basic qualifications includes building and supporting the tools, scripts, and processes for managing data and training pipelines for machine learning.

What You Will Do:

  • Extracting, processing, and organizing raw data (images, lidar).

  • Support for data exploration and visualization tools.

  • ML model training, evaluation, and deployment.

  • Developing and deploying a robust full-stack pipeline that can support various perception/planning/prediction projects.

  • Design and implement scalable and reliable systems for ingestion, processing, and analysis of large disparate data sets from diverse sources.

  • Improve existing and create new data infrastructure components to better automate extraction, transformation, loading, and other data management processes.

  • Develop tools and applications to proactively measure, monitor, and improve data quality and consistency during loading and analysis processes.

  • Analyze and improve efficiency, reliability, and scalability of data infrastructure and processes.

  • Work with data team to define and promote best practices for data management and analysis, and to build and improve systems to implement and support these practices.

Requirements

  • Compares and contrasts the latest developments and emerging issues in the industry.

  • Raises coworkers’ awareness of industry standards, practices and guidelines.

  • Assesses how regulatory and reporting requirements apply to own organization.

  • Strong skills and experience in data schema design and database performance tuning

  • Software Development Life Cycle

  • Describes tasks, tools and practices for converting software product requirements into design.

  • Demonstrates experience with all phases and deliverables of the product development methodology.

  • Assesses the impact of new productivity improvement tools on one’s own area of responsibility.

  • Ability to use agentic workflows to support software development life cycle

  • Application Development Tools

  • Evaluates toolkits used to support major production systems.

  • Experience developing and managing applications on cloud-based platforms such as AWS.

  • Experience in Linux environments

  • Professional experience with ETL, RDBMS and No SQL database.

  • Technical Troubleshooting:

  • Emphasizes the business impact of failure and the criticality and timing of needed resolution so that problems can be avoided in the future.

  • Create trouble reports for all issues found and reviews solutions for completeness and correctness.

  • Directs the resolution of communications problems in multi-vendor environments.

  • Coaches others on advanced diagnostic techniques and tools for unusual or performance-related problems

  • Degree Requirement: Bachelor’s in computer science, Engineering, Mathematics, or an equivalent discipline

Consideration for Top Candidates:

  • 3+ years experience in ML Pipeline Development & Maintenance

Benefits & conditions

Subject to plan eligibility, terms, and guidelines. This is a summary list of benefits.

  • Medical, dental, and vision benefits*

  • Paid time off plan (Vacation, Holidays, Volunteer, etc.)*

  • 401(k) savings plans*

About the company

Your Work Shapes the World at Caterpillar Inc.

When you join Caterpillar, you’re joining a global team who cares not just about the work we do - but also about each other. We are the makers, problem solvers, and future world builders who are creating stronger, more sustainable communities. We don’t just talk about progress and innovation here - we make it happen, with our customers, where we work and live. Together, we are building a better world, so we can all enjoy living in it.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva ¡ JS Congress

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy ¡ LIVE

4:09 min

Challenges of interpreting raw data with language models

Clemens Vasters Clemens Vasters ¡ WWC 2025

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 ¡ LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes ¡ LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou ¡ Coffee With Developers

Videos

See all

Related articles

See all