Senior SRE Engineer
Summary
Senior SRE role focused 70-90% on Infrastructure as Code, automation, and coding for AWS cloud and Kubernetes (EKS) platforms, including AI workloads. Uses Terraform, Python/Go, with emphasis on observability and operational excellence.
Are you an experienced SRE who thrives on automation, Infrastructure as Code, and building highly reliable cloud platforms? We're looking for a Senior Site Reliability Engineer to take ownership of the AWS cloud infrastructure and Kubernetes environment
This is a highly hands-on role, with 70-90% of your time focused on Infrastructure as Code, automation, coding and solution design. You'll work closely with technical leadership, contribute to platform strategy, and help drive engineering excellence across the business.
What You'll Bring
This is a highly hands-on role, with 70-90% of your time focused on Infrastructure as Code, automation, coding and solution design. You'll work closely with technical leadership, contribute to platform strategy, and help drive engineering excellence across the business.
What You'll Bring
- Strong Site Reliability Engineering (SRE) background
- Proven Solid experience designing, maintaining and supporting Kubernetes clusters
- Experience building AI workloads and agents on Kubernetes EKS
- Deep experience with Infrastructure as Code (Terraform)
- Security mindset
- Strong automation mindset
- Development capability using Python and/or Go
- Experience designing scalable, secure and resilient cloud solutions
- Strong understanding of observability, reliability, CI/CD and operational excellence
- Ability to provide technical thought leadership
- AWS Cloud experience with agnostic mindset, with experience applying best practices across cloud platforms