SRE DevOps Engineer

AGM Tech Solutions, LLC
United States
1 day ago

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Artificial Intelligence Amazon Web Services Microsoft Azure Cursor (Graphical User Interface Elements) Linux DevOps Programming Tools Domain Name System (DNS) Revision Control Systems HAProxy
+35 more
Monitoring of Systems Python (Programming Language) Linux System Administration Networking Basics Routing Reliability Engineering Software Tools Prometheus Reverse Proxy Software Engineering TCP/IP Datadog Scripting Google Cloud Enterprise Software Applications Load Balancing GitHub Copilot DevOps Tools - Open-source Grafana Firewalls (Computer Science) Infrastructure as Code (IaC) Git Build Management Kubernetes Infrastructure Automation Frameworks Information Technology Deployment Automation Build Process Terraform Splunk GPT Dynatrace Big Ip Docker Jenkins

Job description

We are seeking a highly technical, hands-on Senior Site Reliability Engineer (SRE) / DevOps Engineer to join our engineering organization. This individual will play a critical role in designing, automating, and maintaining our development and production environments while partnering closely with software engineering teams to improve application reliability, scalability, and deployment efficiency., * Design, implement, and maintain CI/CD pipelines using Jenkins and related DevOps tooling.

  • Partner closely with software engineers to improve build processes, deployment automation, and application reliability.
  • Develop automation scripts and tooling using Python.
  • Troubleshoot complex application, infrastructure, and deployment issues across development, test, and production environments.
  • Manage and optimize Linux-based infrastructure supporting enterprise applications.
  • Configure and support load balancing technologies including F5 BIG-IP and HAProxy.
  • Monitor system health, performance, and availability while proactively identifying opportunities for automation and optimization.
  • Support infrastructure modernization initiatives and implement Infrastructure as Code (IaC) best practices.
  • Collaborate with networking, infrastructure, and application development teams to ensure highly available and scalable solutions.
  • Utilize AI-assisted engineering tools (GitHub Copilot, Cursor, ChatGPT, Claude Code, or similar) to improve development productivity, troubleshooting, documentation, and automation.
  • Participate in production support, incident response, root cause analysis, and continuous improvement initiatives.

Requirements

The ideal candidate has a strong software engineering mindset combined with deep infrastructure expertise. They should understand how applications are built and deployed, be passionate about automation, and have experience leveraging modern AI-powered development tools to accelerate engineering workflows., * 8+ years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, or Infrastructure Engineering.

  • Strong Python scripting and automation experience.
  • Extensive experience with Jenkins and CI/CD pipeline development.
  • Deep understanding of software development lifecycles, application build processes, and deployment methodologies.
  • Experience supporting Java or other enterprise application environments.
  • Strong Linux systems administration experience.
  • Experience with source control systems such as Git.
  • Strong troubleshooting skills across infrastructure, networking, and applications.
  • Experience working in Agile software development environments.

Infrastructure & Networking Experience

  • Strong understanding of networking fundamentals, including:
  • TCP/IP
  • DNS
  • HTTP/HTTPS
  • SSL/TLS
  • Routing
  • Firewalls
  • Load balancing o F5 BIG-IP o HAProxy o Reverse proxy technologies

  • Experience supporting highly available production environments., * Kubernetes and container orchestration experience.
  • Docker.
  • Terraform or other Infrastructure as Code tools.
  • AWS, Azure, or Google Cloud Platform cloud experience.
  • Monitoring and observability tools such as Splunk, Prometheus, Grafana, Datadog, or Dynatrace.
  • Experience implementing SRE best practices, reliability engineering, and production automation.

Preferred Characteristics

  • Extremely hands-on technical contributor.
  • Strong collaboration and communication skills.
  • Automation-first mindset.
  • Passion for continuous improvement and engineering excellence.
  • Comfortable working across development, infrastructure, networking, and operations teams.
  • Curious about emerging technologies and experienced using AI-powered engineering tools to improve productivity and software delivery.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

Videos

See all

Related articles

See all