Site Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+18 more
Job description
- Providing leadership in the Incident resolution process, including creating and maintaining documentation, and providing key input to Post-mortem analysis
- Improving Service Requests and Change Management processes, both technically and through stakeholder management).
- Participate in the process for, and Proactively mitigate risks in a Security management process (Vulnerabilities in Code, Infrastructure, Dependencies)
- Lead discussion in client-facing meetings and discussions around the SRE process, and identifying areas for increasing SRE footprint.
- Engaging with suppliers and 3rd parties for support, requests and opportunities
Requirements
We are seeking an experienced Site Reliability Engineer (SRE) with expertise in Infrastructure as Code tools like Terraform, core CI/CD tools such as Azure DevOps, and monitoring tools including DataDog and AWS CloudWatch. The ideal candidate will have commercial experience in technologies like Dotnet or Java, and be skilled in troubleshooting, incident resolution, and improving service and change management processes. Strong leadership in client-facing discussions and engagement with third-party suppliers is essential. An SRE Foundation certificate and a cloud provider associate-level certification are highly beneficial.
- Commercial experience and proficiency with industry standard:
- IAC tooling (Terraform preferably, or ARM/bicep and CloudFront)
- Core CI/CD Tooling (Azure DevOps, GitHub Actions or Gitlab)
- Monitoring Tooling (DataDog, Splunk, NewRelic, Azure Monitor, AWS CloudWatch)
- Commercial experience in at least one core technology (Dotnet, Java, AI/Data Engineering, Golang)
- Troubleshooting issues and identifying systemic failings indicated by incidents/failures
- Implementing fixes
- Proposing solutions for reducing toil, * 3-9 Years experience
- Bachelor’s degree (or equivalent) in computer science or related discipline
- SRE Foundation certificate (DevOps Institute) and a Cloud provider (AWS, Azure, GCP) ‘associate’-level certification, or completed during the probationary period.
- Proficiency in Azure and Kubernetes, with hands-on experience in managing and deploying applications.
- Expertise in Infrastructure as Code (IaC) using Terraform for efficient and scalable infrastructure management., * Certified Kubernetes Administrator / Application Developer
- Certified Azure DevOps Engineer
- Experience with monitoring tools such as NewRelic or Splunk for effective system monitoring and alerting.
- Familiarity with Harness for continuous delivery and deployment processes.
- Strong programming skills in .Net, Java, or JavaScript for developing robust and scalable applications
About the company
Ensono is a global technology services provider dedicated to helping organizations navigate the complexity of digital transformation. Through Ensono Product, Consulting & Technology, our dedicated consulting arm, we partner with clients to design, build, and modernize digital capabilities across application development and modernization, data platforms and AI, and identity and access management. As we continue to expand our global consulting footprint, we are committed to delivering innovative, high-quality outcomes that enable our clients to move faster, smarter, and more securely.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on dejobs.orgGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Is Software Engineering Over-Saturated?
Find a Developer Job: 12 Best Job Sites For Developers
Where To Find Software Engineering Jobs
Where to Find Entry-Level Software Engineering Jobs