Site Reliability Engineer - Comcast Technology Solutions
Comcast
Charing Cross, United Kingdom
3 days ago
Role details
Contract type
Permanent contract Employment type
Full-time (> 32 hours) Working hours
Regular working hours Languages
English Experience level
SeniorJob location
Charing Cross, United Kingdom
Tech stack
Amazon Web Services (AWS)
Azure
Bash
Cloud Computing
Databases
Continuous Integration
Monitoring of Systems
Python
NoSQL
Performance Tuning
Reliability Engineering
Ansible
SQL Databases
Datadog
Scripting (Bash/Python/Go/Ruby)
Google Cloud Platform
React
Database Performance
Containerization
Kubernetes
Infrastructure Automation Frameworks
Performance Monitor
Front End Software Development
Cloud Optimization
Terraform
Splunk
Docker
Job description
- Analyzes and forecasts system capacity requirements to ensure scalability and performance for high-profile events.
- Participate in incident response efforts, conduct post-incident reviews, and implement corrective actions and improvements to monitoring.
- Develop and maintain monitors and alerts across all services.
- Optimize system performance through tuning and configuration adjustments.
- Develop and maintain disaster recovery plans and procedures to ensure business continuity.
- Create and maintain comprehensive documentation for systems, processes, and procedures.
- Monitor and optimize cloud infrastructure to ensure efficient resource utilization.
- Identify opportunities for automation and implement solutions to reduce manual intervention.
- Work closely with cross-functional teams to align on goals and deliver high-quality solutions.
- Does not have any direct supervisory responsibilities. May direct workflow and act as a technical lead.
- Consistently exercises independent judgment and discretion in matters of significance.
- Shows regular, consistent and punctual attendance.
- Other duties and responsibilities as assigned.
Requirements
Our people are the most important part of our business. We are fundamentally looking for forward-thinking, enthusiastic problem solvers. People who love a challenge, constantly evaluate and question, and, above all, love to ship a product that solves real problems. While these characteristics outweigh any specific technical skills, you should be able to demonstrate some of the following
- An understanding of wider operational performance factors influenced by the underlying infrastructure workload, such as server platforms, databases and networking.
- A strong drive to be a 'detective' and understand why things are working (or not working) as they should, in other words, a passion for detail and an investigative nature.
- The ability to proactively diagnose problems using your holistic knowledge-set - and then get busy with coding a permanent fix, rewriting a process or working with third parties to ensure that lessons are learned, and problems never recur.
- A vision of automation as an opportunity to overcome scale challenges, and a flexible approach to technologies.
Must Have Skills:
- Strong knowledge of cloud platforms (e.g., AWS, GCP, Azure).
- Proficiency in scripting languages (e.g., Python, Bash).
- Experience with infrastructure-as-code tools (e.g., Terraform, Ansible).
- Familiarity with containerization and orchestration tools (e.g., Docker, Kubernetes).
- Experience with monitoring tools (e.g., Datadog, Splunk).
- Experience with Database performance monitoring and tuning (e.g., NoSQL, SQL).
- Experience with Kubernetes performance monitoring and tuning.
- Excellent problem-solving skills and attention to detail.
- Effective communication and collaboration skills.
Desirable Skills:
- CI/CD pipeline management
- Cloud Cost Optimization
- Automating deployment processes
- Front-end development (e.g., React), Bachelor's Degree
While possessing the stated degree is preferred, Comcast also may consider applicants who hold some combination of coursework and experience, or who have extensive related professional experience.
About the company
Comcast brings together the best in media and technology. We drive innovation to create the world's best entertainment and online experiences. As a Fortune 50 leader, we set the pace in a variety of innovative and fascinating businesses and create career opportunities across a wide range of locations and disciplines. We are at the forefront of change and move at an amazing pace, thanks to our remarkable people, who bring cutting-edge products and services to life for millions of customers every day. If you share in our passion for teamwork, our vision to revolutionize industries and our goal to lead the future in media and technology, we want you to fast-forward your career at Comcast., Comcast Technology Solutions is a software technology company headquartered in Denver, Colorado, USA. We enable streaming services, TV stations, pay TV operators, content providers, broadband media sites, and mobile businesses to solve their unique media management and video publishing requirements.
Our Cloud Video Platform (CVP), provided as a service, offers a diverse product catalogue. By leveraging Comcast CVP, our customers can securely manage their digital media, publish content to various IP devices, and effectively monetize their distribution directly to consumers. Our proven media management and publishing technology provides a versatile approach to meet each customer's unique business requirements and scales fluidly to support their growth. Our customers include Deutsche Telekom, Viaplay, Fox, Disney, NBC, Paramount+, and many others.
Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless and robust services to our users. We're a distributed team of engineers with diverse skillsets who thrive on solving complex challenges and driving innovation with a focus on improving observability and reducing toil.