Senior Site Reliability Engineer

Insight Global
Austin, TX, United States
1 day ago
Apply on www.austinjobsite.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Microsoft Access Amazon Web Services Microsoft Azure Cloud Computing Program Optimization Continuous Integration Software Debugging DevOps Distributed Systems IP Routing Reliability Engineering Cloud Services
+9 more
Ansible SAP Business Suiteing Data Logging Infrastructure as Code (IaC) Gitlab Gitlab-ci Kubernetes Terraform Docker

Job description

Our Cloud Infrastructure DevOps team plays a pivotal role in designing and deploying robust infrastructure solutions to support SAP products and services. We are an automation-first organization, prioritizing the deployment of all our cloud resources via automation to enhance efficiency and consistency. Our engineers collaborate closely with internal product teams and customer-facing units to deliver adaptable and scalable code that meets the requirements of our diverse clientele. Candidates will have the opportunity to work across various cloud providers using a wide range of Infrastructure as Code (IaC) tooling and practices, developing provider-agnostic solutions that ensure seamless functionality across different platforms. Our infrastructure supports a broad client base, including local, state, and federal government agencies, as well as private sector organizations engaged in government-related missions., Write, modify, run terraform from an existing codebase to deploy and maintain infrastructure across multiple cloud service providers. Be able to debug errors when deploying terraform.

-Run ansible playbooks to manage customer infrastructure. Be able to modify and troubleshoot ansible as needed as errors occur.

-Use GitLab with multiple repositories to maintain customer infrastructure and create merge requests for changes to customer infrastructure.

-Configure, build, and deploy containerized services using Docker and/or Kubernetes.

-Access traffic flow data between customer and hosted environments to troubleshoot connectivity issues.

-Produce and maintain technical documentation in regard to network and system design and governance.

-Develop standard operating procedures, knowledge base articles, technical bulletins, and other documents in support of the infrastructure.

-Operate in a security-first mindset, performing all other responsibilities with security in mind.

-Implement monitoring, config management, and logging capabilities to manage a multiple tenant cloud infrastructure across multiple cloud service providers.

-Use generative AI elements to increase efficiency and speed, improve accuracy and consistency, enhance security, and better manage resources where practical and within security boundary guidelines.

Requirements

  • 6+ years of experience in a DevOps Engineer or Site Reliability Engineer
  • 5+ years of hands-on Terraform experience for building and maintaining production-grade infrastructure at scale or knowledge of Ansible for automation within the cloud infrastructure
  • Basic cloud networking, example - route tables and security groups
  • Experience working in an AWS and/or either Azure or GCP environment
  • Experience with CI/CD tooling such as Gitlab runners or GitLab CI
  • Experience working in a large enterprise environment with at least 100-200 servers Preferred qualifications:

-Expertise in designing, analyzing and troubleshooting large-scale distributed systems.

-Ability to debug and optimize code and automate routine tasks.

-Systematic problem-solving approach coupled with strong communication skills and a sense of ownership and drive.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.austinjobsite.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all