ironquill.tech/board

$ cat jobs/director-site-reliability-engineering-clover-health-4d019a8d4e8f.json

Director, Site Reliability Engineering

Clover Health·US·Remote - USA·mid
sre
Apply on greenhouse → Get AI match score →
At Counterpart Health, we are transforming healthcare and improving patient care with our innovative primary care tool, Counterpart Assistant. By supporting Primary Care Physicians (PCPs), we deliver improved outcomes at lower cost through early diagnosis and longitudinal care management of chronic conditions. We're looking for a Senior Manager of Site Reliability Engineering to join our team. You'll lead a team of ~10 SREs across North America, UK, HK, and New Zealand — owning both the day-to-day operations and the long-term technical direction of the SRE organization. This role sits at the intersection of people leadership, technical depth, and strategic partnership: you're here to make Counterpart’s infrastructure reliable, scalable, and cost-efficient — and to transform the SRE team's engagement model from reactive support to proactive collaboration with our product engineering pillars. As a Senior Manager, Site Reliability Engineering, you will: Lead and grow our SRE team of ~10 engineers, including hiring, retention, career development, and performance management across multiple time zones (US, HK, NZ). Build strategic partnerships with product engineering pillars — shifting SRE from reactive, ticket-based support to proactive co-ownership of reliability outcomes. Scale our multi-tenant infrastructure to support new customer onboarding and growing patient populations. Own cloud cost management and FinOps practices, building frameworks that balance cost control with reliability and performance. Champion developer self-service and platform engineering. Build self-service capabilities so product teams can manage routine operations without filing SRE tickets. Establish SLOs/SLIs for critical services and improve alert quality so every page is meaningful. Ensure the SRE team is fully leveraging AI tooling in their workflows — using tools like Claude Code for IaC generation, log analysis, root cause investigation, and automating repetitive work — at the same level a

Similar remote roles

Senior Ruby on Rails Engineer
fifth third bank · UK · senior
Site Reliability Engineer - Dedicated Hosted Runners
GitLab · APAC · mid
Site Reliability Engineer (SRE/ DevOps) - Engineering Productivity
Aristanetworks · Worldwide · mid
Executive Director, AI Infrastructure & Platform Engineering
lifelancer · US · mid
Consultant Ingénieur Avant-Vente & Architecte Technique - F/H/N
Octotechnology · Worldwide · mid
Data Engineer
SumUp · EU · mid
Senior Site Reliability Engineer, Workforce Identity
Coinbase · Worldwide · senior
Senior SRE/DevOps Engineer (Remote, Colombia/ Brazil) - Min.
exceptionly · US · senior