System Engineer, Managed Operations
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+15 more
Job description
AWS has successfully launched the European Sovereign Cloud (ESC), marking a significant development in Utility Computing (UC). To spearhead this initiative, we are actively seeking experienced systems engineers with a strong background in cloud operations. As part of the AWS Managed Operations team, you will play a pivotal role in building, operating and evolving operations and development teams dedicated to delivering high-availability AWS services, including EC2, S3, Dynamo, Lambda, and Bedrock, exclusively for EU customers. For more information on ESC please check out our blog: https://aws.amazon.com/blogs/aws/in-the-works-aws-european-sovereign-cloud/
Your responsibilities will encompass overseeing the ongoing operations and expansion of the ESC, working closely with global AWS teams, and influencing the evolution of AWS services and technology. A typical day in this role involves collaborating with technology leaders, contributing to the enhancement of day-to-day operations, and ensuring continuous improvements in availability, reliability, latency, performance, and efficiency of the ESC.
The overarching goal is to deliver scalable services and ensure a high-availability experience for EU customers. If you are an experienced professional ready for a challenging and impactful opportunity, we invite you to join our efforts in operating and scaling a best-in-class development engineering and operations team that aligns with AWS’ commitment to customer satisfaction and continual innovation., You’ll spend a majority of your time operating and improving one of the largest software systems. Over the course of a week, you will review the operational health of the services in your team’s care, and as soon as you figure out why there was an anomaly, you write up an actionable bug report. As a responsible engineer, you’ve learned never to make changes to production systems without a plan, so you reviewed then executed changes following a change management process to one of the production systems in your care. Later in the week, you help to resolve your team’s backlog of operational issues. You round off the week by writing a cool script that you shared with your team which helps get to root cause faster of a hard problem that you diagnosed earlier.
Requirements
Fluency in written and spoken English is required.
Candidate must be a national of an EU member state., * Experience in Linux OS and network troubleshooting, or experience in networking administration and troubleshooting
- Experience in Python, Perl, or another scripting language
- This role requires you to be a national of an EU member state, * Able to execute standard operating procedures and following operational best practices
- Experience in Systems engineering, site reliability engineering, building and operating systems at scale
- Knowledge or experience operating 24x7 high-availability, distributed software applications, performance tuning software applications and optimizing fleet utilization and monitoring frameworks (such as CloudWatch, Datadog, Grafana, Elastic or similar)
- Knowledge or experience with Infrastructure as Code, (such as CDK, CloudFormation, Puppet, Chef, Ansible, or similar)
- Experience with CI/CD pipelines, DevOps practices, and Generative AI technologies, including automated deployment, configuration management, continuous integration workflows, prompt engineering, model deployment, and AI-powered automation tools
About the company
European Sovereign Cloud (ESC) is a part of AWS Utility Computing (UC).
AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon’s Simple Storage Service (S3) and Amazon Elastic Compute Cloud (EC2), to consistently released new product innovations that continue to set AWS’s services and features apart in the industry. As a member of the UC organization, you’ll support the development and management of Compute, Database, Storage, Internet of Things (Iot), Platform, and Productivity Apps services in AWS. Within AWS UC, Managed Operations engineers engage with AWS customers who require specialized security solutions for their cloud services., Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses., AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger, more collaborative teams. Our continual innovation is fueled by the bold ideas, fresh perspectives, and passionate voices our teams bring to everything we do., Amazon is an equal opportunities employer. We believe passionately that employing a diverse workforce is central to our success. We make recruiting decisions based on your experience and skills. We value your passion to discover, invent, simplify and build. Protecting your privacy and the security of your data is a longstanding top priority for Amazon. Please consult our Privacy Notice (https://www.amazon.jobs/en/privacy_page) to know more about how we collect, use and transfer the personal data of our candidates.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Top-Paying Tech Jobs (with Salaries)
Fully Remote Software Engineer Jobs
The Most Popular IT Jobs on the Market
Where To Find Software Engineering Jobs