At a glance
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do
Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time Requisition Job Description Position Summary:
Our Site Reliability Engineering team is the execution engine behind the reliability, availability, and performance of distributed store technology powering thousands of retail and pharmacy locations nationwide
We operate across pharmacy platforms, Point of Sale (POS) systems, handheld devices, store servers, dispensing systems, and edge computing infrastructure — spanning hybrid cloud and on-premises environments at massive fleet scale
Our engineering philosophy is grounded in five pillars: Detection , Prevention , Recovery , Learning Loops , and Developer Experience (DevX) Our operating principle is the reliability covenant : our success is not measured by incident response volume — it is measured by the reliability capability we transfer to the engineering teams we serve
Your success in this role is measured by what the engineering teams in your domain can do independently after working with you , not by how indispensable you become to them
An SSE who has enabled a development team to detect, respond to, and learn from production failures without SRE involvement has delivered the highest-value outcome this role can produce
We track operational toil as an engineering metric Engineers at this level are expected to identify recurring manual work, eliminate it through automation, document the reduction, and treat toil accumulation as a reliability risk — not as a sign of operational expertise.
As a Senior Software Engineer — SRE , you independently own the reliability posture of an assigned engineering domain You are not waiting for direction — you are setting it for your domain
You design the alerting strategy, own the SLO health, lead incident command for production issues, facilitate postmortems, tune anomaly detection models, and partner directly with engineering domain owners to shift reliability left into design
You are a technical mentor to SE-level engineers and an escalation resource during active incidents You have the technical depth to diagnose complex distributed system failures, the data instincts to distinguish genuine anomalies from noise in ML-generated signals, and the organizational skills to drive reliability practice adoption in teams that did not necessarily ask for SRE involvement
Scope: Domain ownership — you operate independently and influence adjacent engineering teams The Environment You Are Joining This role exists inside an active SRE transformation
The domain-based SRE ownership model you will operate within is in its early stages Some of the toolchains you will work with are being built in parallel with the operational work
Engineering domain owners are simultaneously learning what SRE can offer them This is an honest description of the role, not a caveat The opportunity is to shape a domains reliability posture from the ground up in a large-scale, consequential technology environment
The engineering decisions you make will affect pharmacy dispensing, prescription fill workflows, and store operations across thousands of locations If you are energized by the combination of technical depth and organizational building, the scope here is significant
If you are looking for a mature, fully-defined SRE environment where the frameworks and toolchains already exist, this role will feel like a different challenge Success in this role requires patience alongside technical rigor : you will demonstrate value before you demand process change, build credibility before you expect adoption, and earn trust with engineering domain owners through partnership rather than mandate
The operating environment includes an edge computing fleet deployed directly inside store locations — unattended nodes where deployment blast radius is geographic and fleet-wide, not functional and service-scoped
You will develop a fleet operations mindset: the primary failure mode in this environment is deployment and configuration propagation, not service logic.
Experience owning Production Readiness Reviews or service launch gates Hands-on chaos or fault injection experience using LitmusChaos , Chaos Toolkit , or Gremlin TIC (Technical Incident Commander) certification or equivalent structured incident command training Experience operating distributed systems in retail, pharmacy, healthcare, or other operationally sensitive environments where failures have direct patient or customer impact LLM integration for operational use cases (alert summarization, runbook suggestion, incident triage assistance) — design or implementation experience Experience with streaming data platforms: Apache Kafka , Redpanda , Apache Flink , or ksqlDB Familiarity with analytical databases for observability workloads: ClickHouse , Apache Druid , or TimescaleDB Experience with service mesh and traffic management: Istio , Envoy , Linkerd Infrastructure-as-code proficiency at production scale: Terraform , Pulumi , or Ansible What Success Looks Like at 6 Months You own the SLO health and alerting strategy for your assigned domain with full independence — including the ML-based anomaly detection configuration — and the false positive rate is measurably lower than when you joined You have led at least five P1/P2 incident bridges as Incident Commander, with structured postmortems published, action items closed, and at least two incident classes eliminated through root cause remediation rather than symptom suppression You have facilitated your first domain-level PRR and signed off on a service launch At least two engineering teams in your domain have adopted a reliability practice — an SLO health check, a runbook standard, or a deployment gate — that you introduced without top-down mandate You have eliminated measurable toil in your domain: at least three recurring manual tasks automated and documented with before/after comparisons The engineering domain owner you partner with describes SRE as a force multiplier for their teams delivery velocity — not as a gating function that slows them down Education: Bachelors degree in Computer Science, Engineering, or a related field — or equivalent practical experience Anticipated Weekly Hours 40 Time Type Full time Pay Range The typical pay range for this role is: $92,700.00 - $203,940.00 This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls
The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above
Our people fuel our future Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong
Great benefits for great people We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families
This full‑time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well‑being of colleagues and their families
The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility
Additional details about available benefits are provided during the application process and on Benefits Moments We anticipate the application window for this opening will close on: 10/31/2026 Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state and local laws.
CVS Health on Oh My Job
308 open positions right now, including 29 in New Jersey. Average salary across all roles: $23,162–$55,238.
| State | Open positions |
|---|---|
| California | 240 |
| Texas | 159 |
| Maryland | 135 |
| Virginia | 106 |
| New Jersey | 99 |
| New York | 80 |
| Florida | 64 |
| Washington | 60 |
| Illinois | 59 |
| Ohio | 52 |
| Colorado | 39 |
| Massachusetts | 33 |
| Minnesota | 33 |
| Arizona | 30 |
| Alabama | 29 |
| Michigan | 25 |
| Pennsylvania | 24 |
| North Carolina | 21 |
| Arkansas | 20 |
| Georgia | 20 |
| Tennessee | 19 |
| Delaware | 14 |
| Nevada | 13 |
| Missouri | 12 |
| Oklahoma | 12 |
| Indiana | 10 |
| Utah | 10 |
| Kansas | 7 |
| Oregon | 7 |
| Wisconsin | 7 |
| Connecticut | 6 |
| Maine | 6 |
| South Carolina | 5 |
| West Virginia | 5 |
| Hawaii | 4 |
| Iowa | 4 |
| Kentucky | 3 |
| New Mexico | 3 |
| Montana | 2 |
| Nebraska | 2 |
| Idaho | 1 |
Executive Director, Software Engineering- Pharmacy & Consumer Wellness (PCW)
CVS Health • Woonsocket, RI
Senior Software Engineer - SRE, Retail and Pharmacy
CVS Health • Woonsocket, RI
Principal Software Engineer
CVS Health • Work from home, RI
Software Engineer - Sr. Consultant level
VISA • US Austin, TX