Machine Learning Engineer, Alexa AI

Amazon.com, Inc.
Boston, MA, United States
19 days ago

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$143,700.0 - $194,400.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Alexa Data Analysis Software Applications Software Design Patterns Machine Learning Performance Tuning Scrum Methodology Software Deployment Software Engineering Data Processing Chatbots
+6 more
Pytorch Large Language Models Generative AI Information Technology Machine Learning Operations TensorRT

Job description

The Alexa AI team is looking for a passionate, talented, and inventive Machine Learning Engineer with a strong machine learning background, to build capabilities such as fine tuning, distillation, and LLM Inference.

As a ML engineer with the Alexa AI team, you will be responsible for machine learning platform focus on LLM training, production deployment, and optimizations to advance the state of LLMs. You will collaborate closely with Applied Scientists and other MLEs, leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate development of Generative Artificial Intelligence solutions., * Will work with other team engineers to investigate design approaches, prototype new technology and evaluate technical feasibility.

  • Work closely with Applied scientists to process data, scale machine learning models
  • Will work in an Agile/Scrum environment to deliver high quality software.

About the team

Central Analytics and Research Science (CARS) is an analytics, software, and science team within Amazon’s Alexa AI organization. Our mission is to provide scalable and reliable evaluation of the state-of-the-art Conversational AI on how customers perceive the assistants they interact with - from the metrics themselves to software applications to deep dive on those metrics - allowing assistant developers to improve their services.

Requirements

The ideal candidate is passionate about new opportunities and has a demonstrable track record of success in delivering new features and products. A commitment to team work, hustle, and strong communication skills (to both business and technical partners) are absolute requirements. Creating reliable, scalable, and high performance AI products requires exceptional technical expertise, a sound understanding of the fundamentals of Computer Science and Machine Learning. This person has thrived and succeeded in delivering high quality technology products/services in a hyper-growth environment., * 3+ years of non-internship professional software development experience

  • 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience
  • Experience working with PyTorch or JAX software
  • Bachelor’s degree or foreign equivalent in Computer Science, Engineering, Mathematics, or a related field, * Experience working with PyTorch or JAX software, or experience with vLLM, SGLang, TensorRT or similar platforms in production environments
  • Experience developing large model hosting platforms, establishing frameworks, and scaling and optimizing inference system.
  • Experience developing and maintaining MLOps tool in large organizations.

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, MA, Boston - 143,700.00 - 194,400.00 USD annually

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

1:00 min

Introduction to chatbot infrastructure and cloud challenges

Stan Girard Stan Girard · WWC 2024

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

4:30 min

Demonstrating a knowledge graph powered chatbot interface

Tomaz Bratanic · WWC 2023

Videos

See all

Related articles

See all