- Design, develop, and evolve primarily AWS cloud infrastructure as code using Terraform
- Implement GitHub Actions CI/CD pipelines with safeguards to secure deployments
- Administer and evolve multiregion Kubernetes clusters
- Maintain and evolve Helm charts and reusable deployment workflows
- Strengthen observability and feed signals into AI-assisted remediation
- Strengthen infrastructure security through image scanning, policy as code, network segmentation, and firewall configuration
- Collaborate closely with product teams to meet their infrastructure needs
- Enable developers and agents to independently deploy and manage their services
- Participate in the on-call rotation
Requirements
- At least 5 years of experience in DevOps, cloud infrastructure, or software development
- AI-oriented mindset: each of your accomplishments meets an AI-native standard
- Strong experience with cloud platforms (AWS preferred; transferable experience with Azure or GCP is also valued)
- Experience implementing and maintaining CI/CD pipelines and managing deployments
- Proficiency with Kubernetes
- Experience with Docker and containerized applications
- Experience with observability tools
- Excellent collaboration skills and the ability to support development teams
- Strong communicator, able to clearly explain platform complexity and trade-offs to stakeholders
- Fluent English, both spoken and written
Demonstrates expertise in designing and evolving AWS cloud infrastructure using Terraform, implementing CI/CD pipelines with GitHub Actions, and managing Kubernetes clusters. Robust focus on infrastructure security, observability, and collaboration with product teams to enhance deployment processes.