Open role
Staff Site Reliability Engineer
Wonder
New York, NYPosted Sep 29, 2026 · 3h ago$209k – $217k
Full-time$209k – $217kStaffFood
About this role
Join Wonder as a Staff Site Reliability Engineer to architect and maintain resilient, self-healing systems that power student dining experiences across the US. You will own critical production services, AWS infrastructure, and the Kubernetes platform, ensuring scalability and reliability for a rapidly growing customer base. This role offers a unique opportunity to shape incident management, drive observability, and optimize cloud costs in a high-impact environment.
What we are looking for
6- Architect resilient, self-healing systems and co-own critical production services
- Own multi-region resilience, AWS infrastructure as code (Terraform), and the Kubernetes platform (EKS)
- Own the end-to-end observability platform, driving reliability improvements with SLOs
- Design scaling and capacity strategy for highly seasonal traffic
- Shape incident management, lead incident response, and conduct postmortems
- Build and maintain CI/CD pipelines and deployment tooling
Skills mentioned
8PythonAWSKubernetesLinuxMongoDBRedisTerraformCI/CD
