Platform Operations Engineer in Glen Echo
Role details
Job location
Tech stack
Job description
As a Platform Operations Engineer you will work with a team to ensure the availability, reliability, and performance of a full stack, containerized microservices platform. You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas
Requirements
You bring enthusiasm, the ability to work well with people from different disciplines with varying degrees of technical experience, and meet the following qualifications, BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience) with 8+ years of relevant experience; 6+ years with a Master's; additional experience may substitute for a degree
Benefits & conditions
n \n
-
\n System Reliability & Performance - Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian \n
-
\n Monitoring & Observability - Leveraging monitoring tools to proactively detect and resolve issues \n
-
\n Incident Response - Leading triage, troubleshooting, root cause analysis, and post incident reviews \n
-
\n SLIs & SLOs - Defining and tracking reliability metrics \n
-
\n SAFe Agile - Participating in release planning, scrums, design sessions, bug triage, and cross team coordination \n
\n
\n