Systems Development Engineer , Operations Infrastructure Services

Amazon.com, Inc.
Nashville, TN, United States
8 days ago

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$122,800.0 - $166,100.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Agile Methodology Artificial Intelligence C Sharp (Programming Language) C++ (Programming Language) Computer Programming Software Design Patterns Linux Distributed Systems Network Topologies Python (Programming Language) Systems Development Life Cycle
+8 more
Ruby Software Engineering Rust (Programming Language) Mobile Robots Reliability of Systems Build Process Data Pipelines Golang

Job description

Join us in building infrastructure monitoring applications that power Amazon’s global operations. You’ll design and deliver services that process high-volume device telemetry, evaluate network linkages in real time, and provide operators with clear, actionable insights across thousands of sites worldwide.

We’re an agile development team within Operations Infrastructure Services (OIS), part of Amazon Robotics, assisting fulfillment centers, delivery stations, and sortation centers globally. Our monitoring product gives operators a single, reliable view of device and network health-from live infrastructure telemetry to reliance-aware alarming that points to root causes instead of overwhelming teams with duplicate alerts. You’ll have the opportunity to apply modern AI technologies, from AI-assisted incident detection to generative AI tooling that accelerates how we build and operate our network.

You might start your day reviewing a design for integrating metrics from a new device type being onboarded across OIS sites worldwide, then inspect runtime metrics to tune collection thresholds and data quality before shipping a monitoring improvement operators experience immediately. You’ll partner with teammates on design reviews, participate in quick feedback loops, and own features end-to-end - working alongside infrastructure teams to deliver the monitoring experience needed to reduce building downtime and assistance. Throughout the day, you’ll balance designing scalable monitoring solutions with hands-on implementation, working across back-end services and data pipelines while assisting each other’s growth and bringing your authentic perspective to the team., Design, build, and operate scalable monitoring infrastructure on AWS that ingests and processes high-volume device telemetry and network topology data

  • Own systems end-to-end across monitoring tooling and data pipelines, from build through deployment, operation, and ongoing maintenance
  • Partner with infrastructure and operations stakeholders to grasp monitoring gaps and translate them into reliable, automated solutions that reduce manual intervention
  • Apply AI/ML and generative AI techniques to improve detection quality, reduce alarm noise, and streamline operational workflows
  • Raise the bar on system reliability, operational excellence, and automation through thoughtful design, thorough testing, and continuous improvement

A day in the life You might start by reviewing a design for integrating metrics from a new device type being onboarded across all Robotics buildings worldwide, then inspect runtime metrics to tune metric collection and quality before shipping a service improvement our operators experience immediately. Our team values partnership, quick feedback loops, and clear ownership.

Requirements

3+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience

  • Experience in automating, deploying, and supporting large-scale infrastructure
  • Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust
  • Experience with Linux/Unix
  • Experience with CI/CD pipelines build processes

Preferred Qualifications

  • 3+ years of non-internship professional software development experience
  • Experience with distributed systems at scale

Benefits & conditions

Amazon offers a full range of benefits that assist you and eligible family members, including domestic partners. Benefits can vary by location, the number of regularly scheduled hours you work, length of employment, and job status such as seasonal or temporary employment. The benefits that generally apply to regular, full-time employees include:

  1. Medical, Dental, and Vision Coverage
  2. Maternity and Parental Leave Options
  3. Paid Time Off (PTO)
  4. 401(k) Plan

If you are not sure that every qualification on the list above describes you exactly, we’d still love to hear from you! At Amazon, we value people with unique backgrounds, experiences, and skillsets. If you’re passionate about this role and want to make an impact on a global scale, please apply!

About the team We’re a close-knit, agile team that owns infrastructure monitoring for OIS within Amazon Robotics and ships to a worldwide operational fleet. We care deeply about the craft of software and foster each other’s growth. Our vision centers on delivering reliable, intelligent monitoring that helps operators grasp their infrastructure at scale. You’ll work with product and operations stakeholders to turn complex monitoring challenges into simple, elegant solutions. We’re excited to welcome someone who align with our standards for well-architected software and collective problem-solving., The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, TN, Nashville - 122,800.00 - 166,100.00 USD annually

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.amazon.jobs

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

4:30 min

Transitioning from software engineering into a DevOps trajectory

Davide Imola Davide Imola · LIVE

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

Videos

See all

Related articles

See all