About the Role
You will design and implement highly available, scalable systems and tools for engineers. You will build CI/CD workflows, maintain databases, scale web backends, monitor applications, resolve reliability and performance issues, and develop alerts and dashboards for system health.
Requirements
- Experience with AWS cloud infrastructure
- Database management and caching strategy knowledge
- Problem-solving and troubleshooting skills
- Experience with Docker, Kubernetes, and orchestration tools
- Communication and collaboration skills
- Experience with Python and Terraform
- 3+ years of DevOps, SRE, or system administration experience
Responsibilities
- Design and implement highly available, high-performance, and scalable systems
- Design and implement tools for engineers
- Build CI/CD workflows to improve stability and iteration speed
- Maintain and optimize key-value and relational databases
- Scale and load balance web server backends
- Monitor systems and applications and resolve reliability, scalability, and performance issues
- Develop monitoring tools, alerts, and dashboards