About the Role
You will own and improve the end-to-end application delivery platform, from CloudFront through EKS clusters. You will build in-house platform tools and custom Kubernetes controllers, automate operational workflows, maintain delivery infrastructure, and help define reliability standards. You will participate in on-call support, troubleshoot platform issues, and document platform architecture and procedures.
Requirements
- Demonstrable experience with advanced software engineering practices, including unit and integration testing
- Deep understanding of Kubernetes internals and the controller-runtime framework
- Experience developing, testing, and maintaining custom Kubernetes controllers, operators, and CustomResourceDefinitions
- Experience as a Platform Engineer, Site Reliability Engineer, or similar role focused on end-to-end platform ownership
- Practical experience with Karpenter, Argo CD, and Terraform
- Strong networking fundamentals and ability to troubleshoot request paths across CloudFront, WAF, load balancers, Kubernetes networking, service mesh, and workloads
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
Responsibilities
- Develop, maintain, and automate in-house tools, Kubernetes manifest generators, and scripts using Python or Go
- Upgrade and patch EKS clusters and associated components through safe, repeatable automation
- Manage and optimize Karpenter and Argo CD
- Implement and manage service mesh solutions such as Istio and Linkerd
- Participate in a 24/7 on-call rotation and implement preventative measures
- Configure and manage CloudFront distributions and WAF web ACLs
- Develop and maintain platform architecture, process, and troubleshooting documentation
Benefits
- Super Flex Time with no core time
- Annual leave of up to 14 days in the first year
- Personal leave of 5 days each year
- Social insurance
- 401K
- Translation and interpretation support
- Visa sponsorship
- Relocation support