Software Developer/SRE | Berkeley Heights, NJ | ONSITE | W2
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+10 more
Job description
- Monitor production environments to ensure application availability, performance, and reliability.
- Utilize observability and monitoring tools to proactively identify and resolve issues.
- Perform incident management, troubleshooting, root cause analysis, and post-incident reviews.
- Support and maintain Linux and Windows-based environments.
- Manage and troubleshoot containerized workloads running on Kubernetes and AKS.
- Collaborate with development, infrastructure, and operations teams to improve system stability.
- Analyze performance metrics and identify opportunities for process improvement and automation.
- Support applications operating in hybrid cloud and on-premises environments.
- Create and maintain operational documentation, runbooks, and support procedures.
- Participate in production support and reliability initiatives.
Requirements
Do you have experience in Windows support?, * 5+ years of experience in Site Reliability Engineering, Production Support, DevOps, Cloud Operations, or Infrastructure Engineering.
- Strong Linux administration experience.
- Experience supporting Windows Server environments.
- Hands-on experience with Kubernetes and/or Azure Kubernetes Service (AKS).
- Strong experience with Dynatrace and Splunk for monitoring, troubleshooting, and observability.
- Experience with incident management, root cause analysis, and production support.
- Experience working in Azure cloud environments.
- Strong troubleshooting and problem-solving skills.
- Experience supporting applications in hybrid cloud and on-premises environments.
Preferred Qualifications
- Experience with AWS cloud services.
- Experience with Azure DevOps.
- Experience using Ansible for automation and configuration management.
- Experience with Azure Workbooks.
- Terraform experience.
- Knowledge of infrastructure automation and reliability engineering best practices.
Desired Skills
- Site Reliability Engineering (SRE)
- Production Support
- Linux Administration
- Windows Administration
- Kubernetes / AKS
- Dynatrace
- Splunk
- Azure
- AWS
- Azure DevOps
- Ansible
- Terraform
- Monitoring & Observability
- Incident Management
- Root Cause Analysis (RCA)
- Cloud Infrastructure
- Hybrid Environments
What We’re Looking ForWe are looking for a hands-on engineer who thrives in fast-paced production environments and enjoys solving complex infrastructure and application challenges. The ideal candidate will have a strong operational mindset, excellent troubleshooting abilities, and a proven track record of improving system reliability and performance.
Benefits & conditions
Pulled from the full job description
- 401(k)
- Health insurance
- Vision insurance
- Dental insurance, * Health, vision, and dental insurance (single and family coverage)
- 401(k) plan (employee contributions only)
About the company
Experience Matters. Let your experience be driven by our experience. For more than 40 years, Matlen Silver has delivered solutions for complex talent and technology needs to Fortune 500 companies and industry leaders. Led by hard work, honesty, and a trusted team of experts, we can say that Matlen Silver technology has created a solutions experience and legacy of success that is the difference in the way the world works.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Is Software Engineering Over-Saturated?
Where To Find Software Engineering Jobs
Find a Developer Job: 12 Best Job Sites For Developers
What Are The Top Skills Required For Azure Developers?