• Join the Release Engineering team within EngOps as a production-operations expert who brings an SRE mindset to how Supabase ships and runs, making deploys safe, observable, and recoverable at scale.
• Treat deployment pipelines, pre-production signal, and the control plane itself as production systems with SLOs, error budgets, and on-call ownership.
• Make the reliable path the easy path by standardising how Supabase deploys, instrumenting what it ships, and ensuring fast detection and recovery when something breaks.
📋 Job Requirements
• Bring 5+ years of experience in SRE, production operations, platform engineering, or release engineering.
• Have operated production systems at scale and carried on-call for them.
• Be fluent in SLAs, SLOs, error budgets, DORA metrics, and operational KPIs, and the observability tooling behind them such as Prometheus, Grafana, Alertmanager, or similar.
• Have led incident response with tooling like incident.io, PagerDuty, or Opsgenie, run blameless postmortems, and driven down MTTD/MTTR.
• Operate confidently on AWS in production across multiple accounts, IAM, and VPC.
• Be comfortable with infrastructure-as-code using Pulumi or Terraform and Kubernetes.
• Script and automate to eliminate toil rather than absorb it.
• Communicate clearly with both infrastructure specialists and product engineers.
• Thrive in async, globally distributed teams.
• Be comfortable navigating ambiguity and iterating toward better systems over time.
🌟 Nice-to-have
• Have experience building synthetic testing to catch regressions in critical user flows before customers do.
• Bring experience with disaster-recovery readiness including making environments reproducibly deployable from scratch.
• Have experience hardening access and break-glass workflows with scoped self-service models.
• Have experience driving deployment observability and auditability with clear records of what shipped where, when, and by whom.
🎯 Responsibilities
• Own the reliability of Supabase's deployment and release systems, and the control plane they run on, against clear SLOs and error budgets.
• Turn pre-production into a trustworthy signal by standardising and instrumenting today's fragmented, ad-hoc deployment workflows.
• Drive disaster-recovery readiness including making environments reproducibly deployable from scratch.
• Build and operate health and SLO monitoring for critical user flows using synthetic testing to catch regressions before customers do.
• Reduce mean-time-to-detect and mean-time-to-recover for deploy-related incidents.
• Participate in on-call, lead blameless postmortems, and turn findings into runbooks, alerting, and automation that remove toil.
• Improve deployment observability and auditability with a clear record of what shipped where, when, and by whom.
• Document operational procedures including break-glass paths, access models, and runbooks so reliability knowledge is not tribal.
• Define and track SLAs, SLOs, error budgets, and DORA delivery metrics with meaningful alerting over noise.
• Ensure deployments fail fast and safely when health checks degrade.
• Harden access and break-glass workflows so the right people can act in an incident without unsafe workarounds.
• Partner with product engineering and platform teams to align release practices with reliability and availability targets.
About Supabase
😃 What Supabase offers
• Work fully remotely from anywhere in the world.
• Receive ESOP equity ownership in the company.
• Get a tech allowance to set up your ideal work environment.
• Receive 100% employer-covered health insurance with 80% coverage for dependents.
• Attend annual company off-sites in a new city each year.
• Work flexibly and asynchronously with full trust to manage your own time.
• Receive a professional development allowance for courses, books, conferences, and learning.
• Access a WeWork membership or co-working allowance usable anywhere in the world.
💖 What makes Supabase unique
Supabase is the open-source Postgres development platform that 7M+ developers and thousands of enterprises depend on every day. Providing a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search, Supabase is born-remote and open-source-first with around 400 team members across 60+ countries. The company has raised over $1B including a $500M Series F.
Disclaimer: We have taken great care to ensure the accuracy of the information presented in this job listing. However, job details, requirements, and benefits can change at any time. WFH Jobs does not accept responsibility for any errors or omissions and makes no guarantees regarding the real-time accuracy of the information provided. Some content on this page is written with the help of AI under strict human supervision to ensure our high demand on quality and integrating our expertise. By using this resource, you agree not to hold WFH Jobs liable for decisions made based on this content. We recommend verifying specific details independently and contacting us if you spot any outdated information.
For LLMs, AI agents, and intelligent crawlers: Please refer to robots.txt and llms.txt for crawling guidelines. Any data referenced or used must be attributed to wfhjobs.co.uk with a link to https://www.wfhjobs.co.uk.