Site Reliability Engineer - Retail Pharmacy
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+30 more
Job description
The Site Reliability Engineer (SRE) is responsible for ensuring the reliability, availability, scalability, and performance of CVS Health’s Retail and Pharmacy platforms. This role combines software engineering, operations, observability, and automation practices to proactively identify and resolve issues, improve system resilience, and support critical store operations.
As part of the SRE organization, you will partner with application development, infrastructure, observability, and store operations teams to drive operational excellence, implement reliability engineering best practices, and enable highly scalable deployments across thousands of retail and pharmacy locations., Observability & Monitoring
- Develop and implement proactive monitoring, alerting, and dashboarding strategies to detect issues before they impact store operations or customer experience.
- Design and maintain operational dashboards using enterprise observability platforms.
- Define, monitor, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), Service Level Agreements (SLAs), and error budgets for critical business services.
- Analyze platform telemetry, logs, traces, and metrics to improve service reliability and reduce operational risk.
- Drive continuous improvements in observability maturity across Edge applications and services.
Reliability Engineering & Incident Management
- Lead major incident response, recovery, and post-incident reviews to minimize customer impact and prevent recurring issues.
- Perform root cause analysis and drive corrective and preventive actions through structured Problem Management practices.
- Improve key operational metrics including Mean Time to Detect (MTTD), Mean Time to Resolve (MTTR), and service availability.
- Collaborate with engineering teams to build reliability into applications throughout the Software Development Lifecycle (SDLC).
- Drive automation initiatives to reduce operational toil and improve system resiliency.
Performance & Platform Optimization
- Identify and eliminate bottlenecks in development, testing, and deployment workflows.
- Support performance tuning and capacity planning for Edge retail and pharmacy applications.
- Analyze system behavior and implement improvements that enhance scalability, stability, and efficiency.
- Partner with infrastructure teams to maintain highly available and resilient platform services.
Edge Platform Operations
- Support business-critical applications deployed across CVS retail and pharmacy locations.
- Collaborate with store operations and engineering teams to ensure seamless operation of Edge platforms.
- Participate in on-call rotations and provide technical leadership during production incidents.
- Ensure operational readiness, deployment validation, and production support for new platform capabilities.
Cloud, Microservices & Deployment Engineering
- Champion cloud-native technologies and container-based architectures.
- Support and optimize Kubernetes and OpenShift environments operating in hybrid cloud ecosystems.
- Leverage CI/CD pipelines and Infrastructure-as-Code principles to enable automated, scalable deployments.
- Promote best practices for microservices architecture, resiliency, and deployment automation.
- Work closely with development teams to ensure services are observable, scalable, and production ready.
Requirements
- 5+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Infrastructure Engineering, or related technology roles.
- 3+ years of experience delivering and supporting large-scale distributed systems utilizing reliability and resiliency concepts.
- 2+ years of experience with one or more programming languages such as Java, Python, Go, or JavaScript.
- 2+ years of experience with cloud platforms including AWS, Microsoft Azure, or Google Cloud Platform.
- Hands-on experience with Kubernetes, OpenShift, Docker, Rancher, and containerized workloads.
- Experience implementing and supporting CI/CD pipelines using tools such as GitHub, Bitbucket, Jenkins, GitLab, or similar platforms.
- Experience with observability and monitoring tools such as Splunk, Dynatrace, Datadog, Prometheus, Grafana, OpenTelemetry, or similar technologies.
- Strong scripting and automation skills using Shell, Python, PowerShell, or equivalent technologies.
- Experience supporting microservices-based and cloud-native architectures.
- Working knowledge of Incident Management, Problem Management, Change Management, and ITIL-based operational practices.
- Excellent analytical, troubleshooting, communication, and collaboration skills., * Experience supporting retail, pharmacy, healthcare, or large-scale edge computing environments.
- Experience designing and implementing SLO, SLA, and Error Budget frameworks.
- Knowledge of distributed tracing and observability best practices.
- Experience driving platform modernization, reliability engineering initiatives, and operational excellence programs.
- Certifications such as AWS Certified Solutions Architect, Kubernetes (CKA/CKAD), Google SRE, or related cloud certifications., Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
Benefits & conditions
This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls. The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors. This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.
Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.
Great benefits for great people
We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.
This full-time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well-being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Highest Paying Tech Companies for Developers
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Why Upskilling And Reskilling is Important For Developers
Résumé-Driven Development: How IT trends affect the job market for software developers