EngineeringRemote (US) · Full-time
Site Reliability Engineer (SRE)
Keep WUUUH online and trustworthy: deploy pipelines, Postgres health, observability, and graceful degradation when a dependency fails — without turning ops into theater.
Responsibilities
- Own production deploys (Hetzner/systemd release model), rollbacks, and build verification.
- Improve monitoring, alerts, and on-call notes for web + Postgres.
- Harden backups, migrations, and secret handling for country-isolated environments.
- Partner with product on rate limits, paywall gates, and queue health (aid, specialists, escalations).
- Reduce toil: script repeatable checks, document recovery paths.
Requirements
- 3+ years SRE, platform, or backend ops on Linux services and Postgres.
- Strong with systemd, SSH-based deploys, nginx/reverse proxies, and Node.js runtimes.
- Experience with observability (metrics, logs, uptime) and incident response.
- Security-minded defaults: least privilege, no secrets in git, careful migrations.
Nice to have
- Stripe webhook ops, multi-region or multi-DB familiarity, Terraform/Ansible.
- Prior work on member-sensitive or financial-adjacent products.
How to apply
Email careers@wuuuh.com with subject “Application: Site Reliability Engineer (SRE)”, a resume or portfolio link, and a short note on why this role fits. We aim to reply within two weeks.
Email to apply →