> Markdown version of [/jobs/ext/2189075-sr-systems-dev-engineer-amazon-leo-oisl](https://www.wearedevelopers.com/jobs/ext/2189075-sr-systems-dev-engineer-amazon-leo-oisl). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr Systems Dev Engineer, Amazon Leo OISL - **Company:** Amazon.com, Inc. - **Location:** Redmond, WA, United States - **Experience:** Expert - **Salary:** $151,200.0 - $204,600.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Web Services, C++ (Programming Language), Computer Programming, Extract Transform Load (ETL), Distributed Systems, Embedded Software, Apache Hive, Python (Programming Language), Network Troubleshooting, Network Monitoring, Windows PowerShell, Systems Development Life Cycle, Cloud Services, Prometheus, Ruby, Software Engineering, SQL Databases, Rust (Programming Language), Computer Network Operations, Grafana, Apache Spark, Software Troubleshooting, Performance Monitor, Routing & Switching, Data Pipelines, Golang - **Published:** August 22, 2026 - **Apply:** https://www.amazon.jobs/en/jobs/10510452/sr-systems-dev-engineer-amazon-leo-oisl ## About the Role 6+ years of systems design, software development, operations, automation, and process improvement experience - Experience leading the design, automation, deployment, and support of large-scale infrastructure - Experience with PowerShell (preferred), Python, Ruby, or Java - Experience in Network Operations with proficiency in troubleshooting and debugging complex issues across Layer-1 (Optical Networking) or Layer-2/3 (Switching and Routing), 6+ years of experience in satellite network operations, telecommunications, or large-scale distributed network environments - Hands-on experience with network monitoring, availability/performance telemetry, and operational observability (Prometheus, Grafana, or similar) - Experience building data pipelines and analytics tooling (Spark, SQL, ETL workflows) for operational KPI tracking and anomaly detection - Experience troubleshooting Layer-1 and Layer-2/3 network issues in operational environments - Excellent communication skills with ability to influence cross-functional stakeholders and represent operations in engineering discussions - Programming proficiency in Python and at least one additional modern language (Go, Java, C++, Rust), with experience in cloud services (AWS) and distributed systems architecture ## Description We are seeking an exceptional Senior System Development Engineer to serve as the technical lead and senior individual contributor for our Network Operations team within the Optical Inter-Satellite Link (OISL) organization. This is a high-impact role where you will represent the entire Network Operations function of OISL - owning the strategy, tooling, automation, and operational excellence of a revolutionary laser communication network connecting thousands of satellites in space., This is a hybrid lead and senior individual contributor role. You will own the end-to-end Network Operations strategy for OISL, including defining what tools and services need to be built, driving automation of monitoring and incident response, and ensuring operational readiness as the constellation scales. You will be the senior technical voice for OISL Network Operations, partnering across the organization to drive improvements., Network Operations Strategy & Leadership - Own the OISL Network Operations strategy - tooling roadmap, automation priorities, and process maturity - Lead the NetOps on-call program: escalation design, automated ticketing, runbooks, and incident response - Drive operational cadence: metrics reviews, readiness assessments, and capacity planning Automation & Tooling - Build scalable services that automate monitoring, fault detection, classification, and ticketing for the optical inter-satellite link network - Develop data pipelines for KPI tracking, anomaly detection, and proactive issue identification - Build observability platforms - dashboards, alerting, and real-time monitoring at scale - Apply ML/AI for fault correlation, root cause analysis, and predictive failure detection Cross-Functional Collaboration - Partner with hardware, software, and ground network teams to drive system reliability improvements - Lead investigations into complex link failures and constellation-level fault patterns - Represent Network Operations in architecture reviews and program planning Individual Contribution - Hands-on development of tools, services, and automation (Python, Spark SQL, cloud services) - Design testing strategies for distributed satellite network systems - Author technical specifications and operational documentation - Mentor team engineers A day in the life You review the ticket queue and real-time satellite link health, spot degradation patterns, query telemetry, correlate with constellation changes, and engage and work with Hardware/Embedded Software Experts to root cause the issue. You lead cross-functional syncs on recurring failures, propose automated remediation workflows, build automated ticketing pipelines that eliminate manual triage, review teammates' code, update on-call runbooks, and identify observability gaps to close. About the team The OISL Network Operations team is responsible for the operational monitoring, troubleshooting, and reliability of 10,000+ optical laser links carrying customer traffic across a LEO satellite constellation. We build production-grade software services and tools to automate fault detection, classification, and resolution at scale. Our team operates at the intersection of network operations and software development - monitoring laser link health in real time, automating incident workflows, and driving cross-functional improvements with hardware and software engineering teams. ## Related Videos - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Building Systems that Last](https://www.wearedevelopers.com/videos/1389-building-systems-that-last) - [Fireside Chat with Werner Vogels, VP & CTO, Amazon.com & Daniel Gebler, CTO at Picnic](https://www.wearedevelopers.com/videos/1405-fireside-chat-with-werner-vogels-vp-cto-amazon-com-daniel-gebler-cto-at-picnic) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 131 - AI'm not sure about OSS](https://www.wearedevelopers.com/magazine/472-dev-digest-131-ai-m-not-sure-about-oss) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)