Site Reliability Engineer - Retail PharmacyRI - Woonsocket

CVS · RI - Woonsocket

Posted Posted 12 mins ago
Apply on employer site →

Refyne take

This Site Reliability Engineer - Retail Pharmacy role at CVS was posted in the last 5 days. Before applying, run your resume through the checker on the right — most rejections here are keyword and formatting mismatches, not qualifications.

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The Site Reliability Engineer (SRE) is responsible for ensuring the reliability, availability, scalability, and performance of CVS Health's Retail and Pharmacy platforms. This role combines software engineering, operations, observability, and automation practices to proactively identify and resolve issues, improve system resilience, and support critical store operations. As part of the SRE organization, you will partner with application development, infrastructure, observability, and store operations teams to drive operational excellence, implement reliability engineering best practices, and enable highly scalable deployments across thousands of retail and pharmacy locations. ******This position requires working in shifts and on weekends, with compensatory time off provided on weekdays. Key Responsibilities: Observability & Monitoring • Develop and implement proactive monitoring, alerting, and dashboarding strategies to detect issues before they impact store operations or customer experience. • Design and maintain operational dashboards using enterprise observability platforms. • Define, monitor, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), Service Level Agreements (SLAs), and error budgets for critical business services. • Analyze platform telemetry, logs, traces, and metrics to improve service reliability and reduce operational risk. • Drive continuous improvements in observability maturity across Edge applications and services. Reliability Engineering & Incident Management • Lead major incident response, recovery, and post-incident reviews to minimize customer impact and prevent recurring issues. • Perform root cause analysis and drive corrective and preventive actions through structured Problem Management practices. • Improve key operational metrics including Mean Time to Detect (MTTD), Mean Time to Resolve (MTTR), and service availability. • Collaborate with engineering teams to build reliability into applications throughout the Software Development Lifecycle (SDLC). • Drive automation initiatives to reduce operational toil and improve system resiliency. Performance & Platform Optimization • Identify and eliminate bottlenecks in development, testing, and deployment workflows. • Support performance tuning and capacity planning for Edge retail and pharmacy applications. • Analyze system behavior and implement improvements that enhance scalability, stability, and efficiency. • Partner with infrastructure teams to maintain highly available and resilient platform services. Edge Platform Operations • Support business-critical applications deployed across CVS retail and pharmacy locations. • Collaborate with store operations and engineering teams to ensure seamless operation of Edge platforms. • Participate in on-call rotations and provide technical leadership during production incidents. • Ensure operational readiness, deployment validation, and production support for new platform capabilities. Cloud, Microservices & Deployment Engineering • Champion cloud-native technologies and container-based architectures. • Support and optimize Kubernetes and OpenShift environments operating in hybrid cloud ecosystems. • Leverage CI/CD pipelines and Infrastructure-as-Code principles to enable automated, scalable deployments. • Promote best practices for microservices architecture, resiliency, and deployment automation. • Work closely with development teams to ensure services are observable, scalable, and production ready. Required Qualifications • 5+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Infrastructure Engineering, or related technology roles. • 3+ years of experience delivering and supporting large-scale distributed systems utilizing reliability and resiliency concepts. • 2+ years of experience with one or more programming languages such as Java, Python, Go, or JavaScript. • 2+ years of experience with cloud platforms including AWS, Microsoft Azure, or Google Cloud Platform. • Hands-on experience with Kubernetes, OpenShift, Docker, Rancher, and containerized workloads. • Experience implementing and supporting CI/CD pipelines using tools such as GitHub, Bitbucket, Jenkins, GitLab, or similar platforms. • Experience with observability and monitoring tools such as Splunk, Dynatrace, Datadog, Prometheus, Grafana, OpenTelemetry, or similar technologies. • Strong scripting and automation skills using Shell, Python, PowerShell, or equivalent technologies. • Experience supporting microservices-based and cloud-native architectures. • Working knowledge of Incident Management, Problem Management, Change Management, and ITIL-based operational practices. • Excellent analytical, troubleshooting, communication, and collaboration skills. Preferred Qualifications • Experience supporting retail, pharmacy, healthcare, or large-scale edge computing environments. • Experience designing and implementing SLO, SLA, and Error Budget frameworks. • Knowledge of distributed tracing and observability best practices. • Experience driving platform modernization, reliability engineering initiatives, and operational excellence programs. • Certifications such as AWS Certified Solutions Architect, Kubernetes (CKA/CKAD), Google SRE, or related cloud certifications. Education Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience. Anticipated Weekly Hours 40 Time Type Full time Pay Range The typical pay range for this role is: $92,700.00 - $203,940.00 This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls.  The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors.  This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.    Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong. Great benefits for great people We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families. This full‑time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well‑being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility. Additional details about available benefits are provided during the application process and on Benefits Moments . We anticipate the application window for this opening will close on: 10/30/2026 Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state and local laws.

Explore related searches

Browse the wider category, city and company pages this role belongs to.

Similar site reliability engineer jobs

Other openings posted in the last five days that match this role.

More jobs at CVS