
About Supabase
Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth.
About the Role
We're looking for a Release Engineer (SRE) to join our Release Engineering team (part of EngOps). You will be a production-operations expert who brings an SRE mindset to how Supabase ships and runs, making deploys safe, observable, and recoverable at scale.
In this role, you'll treat our deployment pipelines, pre-production signal, and the control plane itself as production systems—with SLOs, error budgets, and on-call ownership. You will make the reliable path the easy path by standardizing deployments, instrumenting releases, and ensuring rapid recovery.
What You'll Be Responsible For
- Reliability Ownership: Own the reliability of deployment and release systems against clear SLOs and error budgets.
- Workflow Standardization: Turn pre-production into a trustworthy signal by standardizing fragmented, ad-hoc deployment workflows.
- Disaster Recovery: Drive readiness by ensuring environments are reproducibly deployable from scratch.
- Observability: Build and operate health and SLO monitoring for critical user flows using synthetic testing.
- Incident Management: Reduce MTTD and MTTR for deploy-related incidents; participate in on-call and lead blameless postmortems.
- Documentation: Maintain clear records of deployments and document operational procedures to eliminate tribal knowledge.
- Operational Excellence: Define DORA metrics, harden access/break-glass workflows, and partner with product engineering to align release practices.
You Might Be a Good Fit If You
- Have 5+ years in SRE, production operations, platform engineering, or release engineering.
- Have operated production systems at scale and carried on-call responsibilities.
- Are fluent in SLAs, SLOs, error budgets, DORA metrics, and observability tooling (Prometheus, Grafana, Alertmanager).
- Have led incident response and driven down MTTD/MTTR.
- Operate confidently on AWS (IAM, VPC) in production.
- Are comfortable with infrastructure-as-code (Pulumi, Terraform) and Kubernetes.
- Thrive in async, globally distributed teams.
What We Offer
- Fully Remote: We hire globally with a co-working allowance.
- ESOP: Equity ownership for every team member.
- Tech Allowance: Budget to set up your ideal work environment.
- Health Benefits: 100% coverage for employees and 80% for dependents.
- Annual Off-Sites: Yearly company-wide gathering in a new city.
- Professional Development: Annual education allowance for learning and growth.
Culture
Async-friendly
Open to
Worldwide
Sign in to track applications and earn points.