Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will maintain reliable infrastructure for large-scale video ingestion and social features. You will own the on-call rotation, lead incident response and postmortems, scale databases, manage infrastructure as code, and work with engineering teams to meet infrastructure needs.
Requirements
- Strong Terraform and infrastructure-as-code experience
- Hands-on Elasticsearch experience for user-facing features
- Deep GCP knowledge, including Kubernetes, VPC, IAM, Cloud Logging, and managed services
- Experience scaling and sharding MySQL or Postgres databases in production
- Incident response and postmortem experience
- Production experience with GitHub Actions
- Clear incident communication skills
- Experience in startup environments
Responsibilities
- Own the on-call rotation
- Lead incident response and communicate during P0 incidents
- Drive actionable postmortems that prevent recurrence
- Scale and shard production relational databases
- Own infrastructure as code at scale
- Work with engineering teams to meet infrastructure needs
- Maintain reliable and scalable infrastructure
Benefits
- Comprehensive medical coverage
- Dental coverage
- Vision coverage
- 401(k)
- Fertility benefits
- Parental benefits
- Generous PTO
- Daily meals
- Wellness benefits
- Commuter benefits at the NYC headquarters
- Learning and development stipend