Site Reliability Engineer III - AWS, Java and Kubernetes
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+15 more
Job description
Experteer Overview In this role you will strengthen the reliability and performance of mission-critical systems within JPMorgan Chase’s Corporate Technology team. You will work on scalable, observable architectures using code and cloud infrastructure, and you will help the team adopt SRE best practices. You’ll contribute to incident triage, post-incident analysis, and proactive improvements tied to SLOs, delivering stable, scalable platforms. This is a hands-on opportunity to influence how we design, deploy, and run reliable systems at scale. Compensation / Benefits * Guide and promote SRE best practices within the team and foster consensus on design decisions * Collaborate with software engineers to design, implement, and test deployment and reliability approaches via CI/CD pipelines * Leverage enterprise AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis while handling data securely * Implement infrastructure, configuration, and network as code for assigned applications and platforms * Work with stakeholders to resolve complex problems and proactively address issues using SLOs/SLIs * Improve availability, reliability, and scalability of applications in collaboration with partners * Identify roadblocks and explore new technologies to solve business problems * Apply SRE fundamentals to monitoring, incident response, capacity planning, and toil reduction Tasks * Formal training or certification in Site Reliability Engineering concepts and 3+ years of applied experience * Exposure to SRE practices for data management/migration platforms; familiarity with on-prem/public cloud components (Compute, Storage, Networks, Database) * Understanding of SRE fundamentals and ability to define and track SLOs/SLIs * Proficiency in at least one programming or configuration tool (e.g., Python, Ansible, Terraform) * Experience with observability and telemetry tools (e.g., Grafana, Dynatrace, Prometheus, Datadog, Splunk) * Experience with public/private/hybrid cloud environments and container orchestration (Kubernetes, ECS, Docker) * Experience with CI/CD tools (Jenkins, GitLab, Terraform) * Familiarity with troubleshooting networking technologies * Experience with containerization and orchestration and related networking issues Key requirements * competitive base salary * incentive compensation may be available * comprehensive health care coverage * retirement savings plan * backup childcare * tuition reimbursement (education assistance)
Requirements
for assigned applications and platforms * Work with stakeholders to resolve complex problems and proactively address issues using SLOs/SLIs * Improve availability, reliability, and scalability of applications in collaboration with partners * Identify roadblocks and explore new technologies to solve business problems * Apply SRE fundamentals to monitoring, incident response, capacity planning, and toil reduction Tasks * Formal training or certification in Site Reliability Engineering concepts and 3+ years of applied experience * Exposure to SRE practices for data management/migration platforms; familiarity with on-prem/public cloud components (Compute, Storage, Networks, Database) * Understanding of SRE fundamentals and ability to define and track SLOs/SLIs * Proficiency in at least one programming or configuration tool (e.g., Python, Ansible, Terraform) * Experience with observability and telemetry tools (e.g., Grafana, Dynatrace, Prometheus, Datadog, Splunk) * Experience with aaaaaaaaaaa Kubernetes cloud environments and container orchestration (Kubernetes, ECS, Docker) * Experience with CI/CD tools (Jenkins, GitLab, Terraform) * Familiarity with troubleshooting networking technologies * Experience with containerization and orchestration and related networking issues Key requirements * competitive base salary * incentive compensation may be available * comprehensive health care coverage * retirement savings plan * backup childcare * tuition reimbursement (education assistance)
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on us.experteer.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Highest Paying Tech Companies for Developers
How Much Does a Software Engineer Make? Realistic Software Engineering Salaries
Software Engineer Salary London
React Developer Salary [2023]