SRE
TSQ SYSTEMS INC
Philadelphia, PA, United States
25 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source
Tech stack
Amazon Web Services
Microsoft Azure
Bash Shell
Cloud Computing
DevOps
Monitoring of Systems
Python (Programming Language)
Network Security
Reliability Engineering
Prometheus
Datadog
Scripting
+12 more
Google Cloud
System Availability
Delivery Pipeline
Grafana
Reliability of Systems
Containerization
Kubernetes
Infrastructure Automation Frameworks
Deployment Automation
Cloudwatch
Terraform
Docker
Job description
As a Site Reliability Engineer (SRE), you will play a crucial role in ensuring the high availability, scalability, and performance of applications and infrastructure in production environments. You will work closely with development and DevOps teams to automate deployment processes, monitor systems, and troubleshoot production issues., * Ensure high availability, scalability, and performance of applications and infrastructure in production environments.
- Monitor systems using tools like Prometheus, Grafana, Datadog, or CloudWatch and respond to incidents.
- Automate deployment, monitoring, and operational tasks using CI/CD pipelines and scripting (Python, Bash, etc.).
- Troubleshoot production issues, perform root cause analysis, and implement preventive measures.
- Collaborate with development and DevOps teams to improve system reliability and release processes.
- Manage cloud infrastructure (AWS, Azure, or Google Cloud Platform) and implement best practices for security and performance.
Requirements
- Experience with monitoring tools such as Grafana.
- Proficiency in scripting languages like Python, Bash, etc.
- Strong problem-solving and troubleshooting skills.
- Knowledge of CI/CD pipelines.
- Experience with cloud platforms like AWS, Azure, or Google Cloud Platform.
Preferred Skills:
- Certifications in relevant technologies.
- Experience with containerization technologies like Docker, Kubernetes.
- Knowledge of networking and security best practices.
- Experience with infrastructure as code tools like Terraform.
- Excellent communication and collaboration skills.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
AJ
Austin Joy
What Are The Top Skills Required For Azure Developers?
over 4 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
about 2 months ago
LM
Luis Minvielle
Is Software Engineering Over-Saturated?
over 2 years ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
over 2 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
over 2 years ago