SRE

TSQ SYSTEMS INC
Philadelphia, PA, United States
25 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing DevOps Monitoring of Systems Python (Programming Language) Network Security Reliability Engineering Prometheus Datadog Scripting
+12 more
Google Cloud System Availability Delivery Pipeline Grafana Reliability of Systems Containerization Kubernetes Infrastructure Automation Frameworks Deployment Automation Cloudwatch Terraform Docker

Job description

As a Site Reliability Engineer (SRE), you will play a crucial role in ensuring the high availability, scalability, and performance of applications and infrastructure in production environments. You will work closely with development and DevOps teams to automate deployment processes, monitor systems, and troubleshoot production issues., * Ensure high availability, scalability, and performance of applications and infrastructure in production environments.

  • Monitor systems using tools like Prometheus, Grafana, Datadog, or CloudWatch and respond to incidents.
  • Automate deployment, monitoring, and operational tasks using CI/CD pipelines and scripting (Python, Bash, etc.).
  • Troubleshoot production issues, perform root cause analysis, and implement preventive measures.
  • Collaborate with development and DevOps teams to improve system reliability and release processes.
  • Manage cloud infrastructure (AWS, Azure, or Google Cloud Platform) and implement best practices for security and performance.

Requirements

  • Experience with monitoring tools such as Grafana.
  • Proficiency in scripting languages like Python, Bash, etc.
  • Strong problem-solving and troubleshooting skills.
  • Knowledge of CI/CD pipelines.
  • Experience with cloud platforms like AWS, Azure, or Google Cloud Platform.

Preferred Skills:

  • Certifications in relevant technologies.
  • Experience with containerization technologies like Docker, Kubernetes.
  • Knowledge of networking and security best practices.
  • Experience with infrastructure as code tools like Terraform.
  • Excellent communication and collaboration skills.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

1:07 min

Architecting the availability stack with Prometheus and Grafana

Gabriel Labachelerie · World Congress 2023

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all