ironquill.tech/board

$ cat jobs/site-reliability-engineer-supabase-1ec53d510fcf.json

Site Reliability Engineer

Supabase·Worldwide·Remote·mid
postgressre
Apply on ashby → Get AI match score →
ABOUT SUPABASE Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. ABOUT THE ROLE Supabase manages millions of Postgres instances and is growing. We have strong teams across observability, release engineering, and incident management — and we're concentrating our reliability efforts into a dedicated SRE practice that ties the discipline together across the platform. You'll be embedded within Service Operations, and your primary job is to make every engineering team more reliable — not by owning their infrastructure, but by establishing the practices, frameworks, and feedback loops that let them own reliability themselves. You'll work across the org: sometimes setting the standard, sometimes pair-programming a fix, sometimes helping a team define their error budget, sometimes telling them it's exhausted. This role is ideal for someone who has a strong vision for how SRE should work and thrives in async, fast-paced environments where influence matters more than authority. WHAT YOU'LL OWN - Partner with service teams to define meaningful SLIs and SLOs grounded in customer experience, and build the error budget policies that turn them into engineering decisions - Own and evolve the Operational Readiness Review (ORR) process — conducting reviews for new services and major changes across observability, alerting, runbooks, capacity, and graceful degradation - Strengthen the incident-to-improvement pipeline: connecting postmortem findings to operational readiness gaps, identifying repeat failure patterns, and driving systemic fixes - Act as the reliability expert teams pull in for architecture reviews, failure mode analysis, dependency mapping, and resilience design - Identify and quantify operational toil across the org, and build or advocate for automation that elim

Similar remote roles

Site Reliability Engineer, Enterprise Technology Services
Apple · Worldwide · mid
Software Engineer, Large-Scale Data Systems
Smartly · Worldwide · mid
Associate Principal Engineer, Dotnet Fullstack
Nagarro1 · Worldwide · senior
Specialist DevOps Engineer, Actimize (AWS, PostgreSQL)
NICE · APAC · mid
Staff Platform Engineer - EU
Ashby · Worldwide · senior
Staff Platform Engineer - EU
Ashby · Worldwide · senior
DevOps Engineer
Tiebreak · Worldwide · mid
Platform Engineer
Chronograph (chronograph.pe) · Worldwide · mid