Staff Software Engineer - AI Platform

General Motors
Austin, TX, United States
2 days ago
Apply on www.austinjobsite.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$189,300.0 - $333,400.0
Working hours
Regular working hours

Tech stack

AI Evaluation Java (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Automation of Tests Microsoft Azure Cloud Computing Cloud Engineering Code Review Continuous Integration Data as a Services Fault Tolerance
+22 more
Identity and Access Management Python (Programming Language) Performance Tuning Cloud Services Azure Machine Learning Search Technologies Software Engineering Systems Integration Strategies of Testing Enterprise Software Applications Chatbots Retrieval-Augmented Generation Large Language Models Multi-Agent Systems Caching Build Management AI Platforms Kubernetes Infrastructure Automation Frameworks Low Latency Production Code Agent2Agent Protocol

Job description

We are looking for a Staff Individual Contributor who combines exceptional software-engineering execution with enterprise architecture leadership to design and build the next generation of customer-facing AI systems.

This role will lead the architecture and hands-on implementation of reusable modules for conversational AI, agent orchestration, grounding, memory, tool integration, policy enforcement, and AI platform services. The successful candidate will help establish durable technical patterns across cloud services and customer experiences while personally contributing production code in Python and Go and/or Java.

This is a senior technical leadership role, not a people-management position. The engineer will influence architecture, mentor through technical leadership, raise engineering standards, and drive execution through design reviews and implementation-not through direct reports.

What You’ll Do:

  • Own architecture and hands-on delivery of complex customer-facing AI modules from PRD through production.
  • Turn product needs into clear component boundaries, APIs, data models, ERDs, sequence flows, deployment designs, test strategies, and AI evaluation plans.
  • Build reusable services for orchestration, agents, grounding, retrieval, memory, tool integration, conversation state, and channel adapters.
  • Lead technical decisions for inference, serving, reliability, security, observability, capacity, and operational readiness.
  • Establish reference implementations and engineering patterns for AI-enabled services across GCP and Azure.
  • Drive CI/CD, automated testing, AI evaluation and regression pipelines, release readiness, and production operations.
  • Partner with product, security, privacy, data, infrastructure, and application teams to resolve cross-system tradeoffs.
  • Provide technical leadership through architecture reviews, design documentation, code reviews, incident analysis, and production-readiness reviews.
  • Identify opportunities to simplify duplicated capabilities and standardize reusable platform interfaces., This role is categorized as hybrid. This means the selected candidate is expected to report to a specific location at least 3 times a week {or other frequency dictated by their manager}.

Requirements

  • 8+ years of software engineering experience, including ownership of distributed, cloud-native, or customer-facing systems.
  • Track record as a hands-on technical architect, staff/principal engineer, or lead programmer delivering production systems at scale.
  • Expert Python and strong Go and/or Java programming skills.
  • Strong command of distributed-system design, service boundaries, API and data contracts, asynchronous processing, eventing, caching, consistency, and fault tolerance.
  • Production experience with LLM applications, agent orchestration, RAG, embeddings/vector search, tool use, MCP, A2A, and chatbot or workflow-based systems, including quality, safety, grounding, and regression evaluation.
  • Experience with inference and serving for low latency, high throughput, concurrency, autoscaling, traffic management, and cost/performance optimization.
  • Cloud and delivery experience with GCP and/or Azure, containers/Kubernetes, IAM, secrets, messaging, managed data services, CI/CD, infrastructure as code, and automated testing.
  • Ability to define and implement architecture artifacts and non-functional requirements covering security, privacy, observability, SLOs/SLIs, capacity, BCP, DR, and operational resilience.
  • Strong written and verbal communication, technical judgment, and ability to influence without direct authority.

What Will Give You A Competitive Edge (Preferred Qualifications):

  • Experience building conversational AI for automotive, mobility, contact center, consumer, or other high-scale customer-facing domains.
  • Experience with Vertex AI, Azure AI services, model gateways, vector databases, retrieval/evaluation platforms, and model observability.
  • Experience with privacy-aware personalization, governed memory, grounding, tool access, auditability, and deletion workflows.
  • Experience integrating CRM, identity, knowledge, telephony, messaging, or other enterprise systems.
  • Experience converging duplicated platforms through incremental adoption and leading cross-organization architecture initiatives without direct authority.

Benefits & conditions

Compensation: The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The salary range for this role is $189,300-$333,400 . The actual base salary a successful candidate will be offered within this range will vary based on factors relevant to the position, as well as the geography of the selected candidate.

Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.

Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation and holidays, tuition assistance programs, employee assistance program, GM vehicle discounts, and more.

About the company

At General Motors, our product teams are redefining mobility. Through a human-centered design process, we create vehicles and experiences that are designed not just to be seen, but to be felt. We’re turning today’s impossible into tomorrow’s standard -from breakthrough hardware and battery systems to intuitive design, intelligent software, and next-generation safety and entertainment features.

Every day, our products move millions of people as we aim to make driving safer, smarter, and more connected, shaping the future of transportation on a global scale, We believe we all must make a choice every day - individually and collectively - to drive meaningful change through our words, our deeds and our culture. Every day, we want every employee to feel they belong to one General Motors team., General Motors is committed to being a workplace that is not only free of unlawful discrimination, but one that genuinely fosters inclusion and belonging. We strongly believe that providing an inclusive workplace creates an environment in which our employees can thrive and develop better products for our customers.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.austinjobsite.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:28 min

Defining critical competencies for automotive AI engineering

Daniel Graff +1 · World Congress 2021

3:15 min

Reversing the caching model for artifact delivery

Thijs Feryn Thijs Feryn · World Congress 2026 Europe

1:00 min

Introduction to chatbot infrastructure and cloud challenges

Stan Girard Stan Girard · World Congress 2024

1:15 min

Deploying local container pods to Kubernetes clusters

Stevan Le Meur Stevan Le Meur · World Congress 2024

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:33 min

Maintaining prompt structures for prefix caching

Douglas Reiser Douglas Reiser · Europe 2026 Virtual

Videos

See all

Related articles

See all