Skip to main content
Supabase

Release Engineer

RemoteWorldwide
Published
Role
SRE
Experience
Senior
Employment
Full-time
Company size
Mid-size
Salary not disclosed
Check eligibility

Open to Worldwide. Set where you work from to check your eligibility.

No BS summary

Release/SRE engineer with 5+ years in SRE, production operations, platform engineering, or release engineering. Needs production AWS, Kubernetes, IaC, observability/SLO tooling, incident response, and on-call experience. Fully remote and hired globally.

Core skills

Prometheus/Grafana/AlertmanagerAWSKubernetes

Required skills

incident.io/PagerDuty/OpsgenieIAMVPCPulumi/Terraform

What you'll do

  • Own the reliability of Supabase's deployment and release systems and the control plane they run on against clear SLOs and error budgets.
  • Turn pre-production into a trustworthy signal by standardizing and instrumenting fragmented, ad-hoc deployment workflows.
  • Drive disaster-recovery readiness, including making environments reproducibly deployable from scratch.
  • Build and operate health and SLO monitoring for critical user flows using synthetic testing to catch regressions before customers do.
  • Reduce mean-time-to-detect and mean-time-to-recover for deploy-related incidents.
  • Participate in on-call.
  • Lead blameless postmortems.
  • Turn incident findings into runbooks, alerting, and automation that remove toil.
  • Improve deployment observability and auditability, including a clear record of what shipped where, when, and by whom.
  • Document operational procedures, including break-glass paths, access models, and runbooks.
  • Define and track SLAs, SLOs, error budgets, and DORA delivery metrics with meaningful alerting over noise.
  • Ensure deployments fail fast and safely when health checks degrade.
  • Harden access and break-glass workflows so the right people can act in an incident without unsafe workarounds.
  • Partner with product engineering and platform teams to align release practices with reliability and availability targets.

What they require

  • 5+ years in SRE, production operations, platform engineering, or release engineering.
  • Experience operating production systems at scale and carrying on-call for them.
  • Fluency in SLAs, SLOs, error budgets, DORA metrics, operational KPIs, and the observability tooling behind them.
  • Experience leading incident response with tooling like incident.io, PagerDuty, or Opsgenie, running blameless postmortems, and driving down MTTD/MTTR.
  • Confident production operation on AWS, including multiple accounts, IAM, and VPC.
  • Comfortable with infrastructure-as-code and Kubernetes.
  • Ability to script and automate to eliminate toil rather than absorb it.
  • Clear communication with both infrastructure specialists and product engineers.
  • Ability to thrive in async, globally distributed teams.
  • Comfortable navigating ambiguity and iterating toward better systems over time.

Benefits

  • Fully remote work.
  • Global hiring from anywhere.
  • WeWork membership or co-working allowance usable anywhere in the world.
  • ESOP equity ownership for every team member.
  • Tech allowance for laptop, monitor, headphones, or other work environment needs.
  • 100% health insurance coverage for employees.
  • 80% health insurance coverage for dependents.
  • Annual company off-sites in a new city.
  • Flexible asynchronous work.
  • Annual professional development education allowance for courses, books, conferences, or other learning.

open source backend platform for app development

TechnologyMid-sizesupabase.com

Details

Apply routeDom
Salary not disclosed