Senior AWS Site Reliability Engineer / Infrastructure Engineer

Insight Global
Roseland, NJ, United States
5 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Cloud Computing Cloud Engineering Configuration Management Continuous Integration Identity and Access Management Python (Programming Language) Reliability Engineering Amazon Simple Notification Service (SNS) Software Deployment
+18 more
Data Processing Software Modules Load Balancing Autoscaling System Availability AWS Lambda Amazon Virtual Private Cloud (VPC) Amazon Relational Database Service Infrastructure Automation Frameworks Information Technology Route53 Functional Programming Cloudwatch Amazon Simple Queue Service (SQS) Terraform Code Restructuring Serverless Computing Jenkins

Job description

This is a hands-on engineering role requiring deep expertise in AWS, Terraform, Jenkins, and Python. The ideal candidate will be capable of independently owning technical initiatives, collaborating with cross-functional stakeholders, and delivering cloud infrastructure solutions with minimal oversight.

Day to Day Responsibilities :

Cloud Infrastructure & Reliability Engineering

  • Design, implement, and maintain scalable, secure, and highly available AWS infrastructure.

  • Support ongoing cloud engineering initiatives and operational activities within a large enterprise environment.

  • Act as a technical lead for infrastructure-related projects, partnering with engineering, security, networking, and application teams.

  • Troubleshoot complex infrastructure, reliability, and performance issues across AWS services.

  • Implement best practices around resiliency, observability, scalability, and operational excellence.

Infrastructure as Code (Terraform)

  • Build new Terraform modules from scratch to support enterprise cloud initiatives.

  • Enhance, refactor, and maintain existing Terraform codebases.

  • Establish reusable infrastructure patterns and standards across AWS environments.

  • Ensure infrastructure deployments are automated, version-controlled, and repeatable.

CI/CD & Automation

  • Design, build, and maintain Jenkins pipelines supporting infrastructure and application deployments.

  • Develop new CI/CD workflows while also optimizing existing pipelines.

  • Drive automation initiatives to reduce manual effort and improve operational efficiency.

  • Collaborate with development teams to streamline deployment and release processes.

Python Development & Infrastructure Automation

  • Develop automation tools and scripts using Python.

  • Build solutions supporting:

  • Infrastructure provisioning and configuration management

  • AWS Lambda functions and serverless automation

  • Monitoring and operational tooling

  • Infrastructure health checks and remediation workflows

  • Data processing and cloud operations automation

  • Create scalable solutions to improve platform reliability and operational efficiency.

Stakeholder Management

  • Partner directly with technical and business stakeholders to gather requirements and deliver solutions.

  • Take ownership of projects from initial design through implementation and production support.

  • Provide technical guidance and recommendations related to AWS cloud best practices.

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global’s Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.

Requirements

  • 10+ years of overall IT infrastructure, cloud engineering, or SRE experience.

  • Extensive hands-on experience with AWS in enterprise environments.

  • Strong expertise across core AWS services such as:

  • EC2, VPC, IAM, S3, RDS, Route 53, CloudWatch, Lambda, SNS/SQS, Auto Scaling, Load Balancing

  • Expert-level Terraform experience, including module development and infrastructure automation.

  • Strong experience designing and maintaining Jenkins CI/CD pipelines.

Advanced Python development skills with experience building automation and operational tooling.

About the company

One of Insight Global’s largest Payroll Clients is seeking a highly experienced AWS Site Reliability Engineer (SRE) / Infrastructure Engineer to join their Public Cloud Engineering team. This individual will play a critical role in supporting and enhancing their AWS cloud infrastructure while driving infrastructure automation, reliability, and operational excellence across enterprise-scale cloud environments.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:23 min

Reviewing AWS infrastructure deployment configuration and planning

Devlin Duldulao · LIVE

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:02 min

Applying an ETL methodology to infrastructure configuration management

Axel Barbier · World Congress 2023

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

57 sec

Extracting API schemas automatically during continuous integration builds

Axel Barbier · World Congress 2023

1:24 min

Evaluating formal AWS certifications versus raw practical engineering experience

Jan Giacomelli · LIVE

Videos

See all

Related articles

See all