Senior Data Engineer - Agentic AI, Automation, and Data Platforms

General Motors
Warren, MI, United States
4 days ago
Apply on generalmotors.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$138,700.0 - $173,750.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Data Analysis Build Automation Automation of Tests Microsoft Azure Continuous Integration Information Engineering Data Integration Data Transformation Data Security
+21 more
Cursor (Graphical User Interface Elements) Distributed Computing Environment Python (Programming Language) Scala (Programming Language) Search Technologies Software Engineering SQL Databases Data Streaming Enterprise Data Management Feature Engineering GitHub Copilot Large Language Models Apache Spark Data Lakes Infrastructure Automation Frameworks Information Technology Data Management Machine Learning Operations Virtual Agents Data Pipelines Databricks

Job description

This role is categorized as hybrid. This means the successful candidate is expected to report to GM Warren Global Technical Center or Austin Technical Center three times per week, at minimum [or other frequency dictated by the business if more than 3 days]., This role is for a senior individual contributor in Data Engineering who can independently lead complex technical work, apply strong professional judgment, improve processes and delivery patterns, and move quickly from ideas to production solutions. At this level, the individual is expected to operate with minimal guidance, resolve non-standard problems using advanced analytical thinking, take ownership of outcomes, and serve as a technical resource for less experienced team members.

The role is anchored in data engineering with a strong focus on automation and Agentic AI. The engineer will build reliable data platforms and use technologies such as Cursor, large language models, Vector Search, Databricks agents, RAG, and similar tools to accelerate engineering delivery and enable intelligent data experiences. The engineer will partner with data scientists and ML engineers as needed to support experimentation and productionize AI solutions, while data engineering and platform delivery remain the primary focus.

What You’ll Do

  • Design, build, and productionize reliable, scalable, and secure data pipelines and data products in Azure Databricks that support AI, analytics, and operational use cases.
  • Transform raw data from multiple source systems into trusted, well-structured data products for analytics, model development, LLM applications, Vector Search, and AI agents.
  • Build automation and reusable engineering workflows using tools such as Cursor, Claude, LLMs, Databricks agents, and related technologies to improve development speed, testing, documentation, troubleshooting, and operational efficiency.
  • Design and enable governed data, retrieval, and semantic patterns for Vector Search, RAG, Databricks agents, Genie, Glean, and other AI-enabled applications.
  • Build and optimize batch and streaming pipelines, including feature-ready, training, inference, and model-scoring data workflows, in partnership with data science teams when needed.
  • Establish practical engineering patterns for CI/CD, automated testing, data quality, lineage, observability, security, cost management, and production support.
  • Solve complex data engineering, performance, reliability, and data-quality problems with strong ownership, urgency, and sound technical judgment.
  • Contribute to technical direction, reusable standards, and delivery practices across teams; influence adoption through working examples and measurable outcomes.
  • Mentor team members through technical guidance, design reviews, knowledge sharing, and strong engineering practices.

Requirements

  • Bachelor’s degree in Computer Science, Software Engineering, Data Engineering, or related field, or equivalent experience.
  • 5+ years of relevant professional experience, or equivalent knowledge and experience.
  • Strong experience in data engineering, including pipeline development, data modeling, data integration, distributed processing, and production support for enterprise data platforms.
  • Experience using Python or Scala, SQL, Apache Spark, and modern cloud data platforms; Azure is preferred, and AWS or GCP experience is also considered.
  • Experience designing, building, and optimizing scalable batch and streaming data pipelines using Databricks, Delta Lake, and medallion or comparable lakehouse architecture.
  • Hands-on experience using AI-assisted development or automation tools such as Cursor, Claude, GitHub Copilot, or comparable platforms to improve engineering productivity and delivery.
  • Hands-on experience with LLMs, Vector Search, RAG, Databricks agents, or comparable technologies used to build or enable production AI solutions.
  • Demonstrated ability to work independently, move quickly through ambiguity, influence technical decisions, and deliver measurable improvements in quality, reliability, efficiency, or business value.

What Can Give You a Competitive Advantage (Preferred Qualifications)

  • Experience building or operating Databricks agents, Vector Search solutions, LLM applications, RAG workflows, Genie spaces, Glean integrations, or similar Agentic AI platforms.
  • Experience applying evaluation, monitoring, access controls, guardrails, and governance to AI or agent-enabled solutions.
  • Experience partnering with data scientists or ML engineers on feature engineering, experimentation, model development, model serving, or productionization of AI solutions.
  • Experience with infrastructure as code, APIs, data contracts, platform automation, or reusable engineering libraries and templates.
  • Experience in manufacturing, supply chain, automotive, planning, or another complex operational domain.
  • Demonstrated mentoring, technical leadership, and process improvement impact consistent with a Level 7 senior individual contributor role.
  • Master’s degree in Computer Science, Software Engineering, Data Engineering, or related field.

Benefits & conditions

  • The expected base compensation for this role is: $138,700 - $173,750. Actual base compensation within the identified range will vary based on factors relevant to the position.
  • Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.
  • Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays, tuition assistance programs, employee assistance program, GM vehicle discounts and more.

About the company

We believe we all must make a choice every day - individually and collectively - to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team., General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on generalmotors.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · World Congress 2026 Europe

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:04 min

Comparing offline data analytics with online stream processing

Artem Volk Artem Volk +1 · World Congress 2024

Videos

See all

Related articles

See all