Senior Site Reliability Engineer

Akamai Technologies
Topeka, United States of America
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior

Job location

Topeka, United States of America

Tech stack

Big Data
Configuration Management
Computer Engineering
Database Queries
Distributed Systems
Monitoring of Systems
Python
PostgreSQL
Metadata
Network Diagnostics
Performance Tuning
Reliability Engineering
Akamai
SQL Databases
Scripting (Bash/Python/Go/Ruby)
Computer Network Operations
Information Technology

Job description

Would you enjoy improving stability and safety of one of the largest global networks?

Would you enjoy hands-on network operations work on a global scale to improve our operational efficiency?

Join Our Metadata Team!

The Metadata Team is a part of the Akamai's Platform Communications group. A cross-functional engineering team that develops the distributed systems and services that underpin Akamai's global network. Our system supports fast and reliable configuration of Akamai's global network. Also, we support the control services for global content management.

Partner with the best

In this role, you'll leverage your big data, distributed systems, network diagnostics and analytical skills to characterize performance, reliability and capacity of the system. You'll help create and maintain the test analysis & simulation tools, supporting performance analysis of the team's various metadata systems.

As a Senior Site Reliability Engineer, you will be responsible for:

  • Tuning systems to optimize performance and to operate more reliably

  • Providing ongoing technical assistance in areas including model database management, configuration management, and simulation runs

  • Managing the rollout and activation of new features and platform changes

  • Developing monitoring tools and automate processes to help scale our systems better

  • Troubleshooting complex application issues, service incidents, performance and availability issues

Requirements

  • Have 5 years of relevant experience and a Bachelor's degree in Computer Engineering, Computer Science or equivalent

  • Have extensive experience with Linux/Unix operating systems and scripting, ideally with Python

  • Have a significant background in performance analytics and performance optimization

  • Be able to monitor systems and applications for performance

  • Have experience with database queries or big data technologies (Postgres/SQL)

  • Have good attention to detail and excellent problem solving/troubleshooting skill

About the company

At Akamai, we make life better for billions of people, trillions of times a day. Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to he

Apply for this position