Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own the on-call rotation, lead incident response and postmortems, and partner with engineers on infrastructure needs. You will scale infrastructure and databases, implement durable fixes, and improve reliability for video ingestion and social features.
Requirements
- Strong Terraform infrastructure-as-code experience at scale
- Hands-on Elasticsearch experience for user-facing features
- GCP experience, including Kubernetes, VPC, IAM, Cloud Logging, and managed services
- Experience scaling and sharding MySQL or PostgreSQL databases in production
- Incident response and postmortem experience
- GitHub Actions CI/CD experience in production
- Experience in startups or rapid-growth environments
Responsibilities
- Own the on-call rotation
- Lead incident response
- Drive actionable postmortems that prevent recurrence
- Work with engineering teams to meet infrastructure needs
- Scale infrastructure and relational databases
- Implement durable fixes for reliability issues
- Communicate issues rapidly during incidents
Benefits
- Comprehensive medical, dental, and vision coverage
- 401(k)
- Fertility and parental benefits
- Generous PTO
- Daily meals
- Wellness benefits
- Commuter benefits at the NYC headquarters
- Learning and development stipend
- Equity