
Principal Cloud Kubernetes Engineer (AWS & Azure)
Skills
Location
Languages
Job Description
We're looking for a Principal Cloud Kubernetes Engineer (AWS & Azure) to design, deploy, and operate multi-cloud Kubernetes infrastructure using GitOps and infrastructure as code to enable scalable, reliable, and secure platform operations.
🎯 Responsibilities
- Design, implement, and maintain production Kubernetes clusters on EKS and AKS across multiple environments and regions
- Configure and operate GitOps workflows with ArgoCD for declarative application delivery and cluster configuration
- Develop and manage infrastructure as code using Terraform for cloud resources and Helm/Helmfile for Kubernetes deployments
- Build observability and reliability practices using Prometheus, Grafana, and CloudWatch
- Implement monitoring and alerting strategies
- Apply security best practices, including network policies, RBAC, pod security controls, and secrets management
- Operate and optimize Kubernetes add-ons such as Argo Workflows, Argo Rollouts, cert-manager, CSI drivers, and cluster autoscalers
- Diagnose and resolve distributed systems issues across compute, networking, and storage layers
- Collaborate with development teams to improve deployment models, scalability, and resource utilization
- Document architectures, operational workflows, and incident-handling standards
🛠️ Requirements
- At least 7 years of professional engineering experience with modern cloud-native architectures
- Strong expertise in AWS and Azure, including EKS, AKS, VPC/VNet, IAM, EBS/EFS, S3, load balancers, and Route 53
- Deep understanding of Kubernetes architecture, networking, storage, and security principles, with experience managing production-scale environments
- Proven experience with infrastructure-as-code workflows using Terraform, including module development and state handling
- Strong experience with Helm and Helmfile
- Practical knowledge of GitOps tools such as ArgoCD or Flux and declarative infrastructure deployment
- Experience implementing security policies, secrets management, and compliance automation
- Advanced knowledge of observability tools, including Prometheus and Grafana, and log management with CloudWatch or ELK/EFK
- Scripting proficiency in Bash, Python, or Go for automation and systems tooling
- Strong knowledge of containerization and image lifecycle management using Docker, ECR, or ACR
- Clear communication skills and the ability to document complex systems effectively
➕ Nice to have
- Familiarity with policy-as-code frameworks such as Kyverno or OPA/Gatekeeper
- Experience with AWS PrivateLink, VPC peering, and hybrid or multi-cloud networking
- Understanding of service mesh patterns, ingress controllers, and DNS management
- Experience with performance tuning and capacity planning for Kubernetes environments
🤗 Benefits
- Private health insurance
- Employee stock purchase plan
- 100% paid sick leave
- Referral program
- Professional certification support
- Language courses
- 24 working days of annual leave
- Paid time off for public holidays
- Career development planning, internal training, mentorship, sponsored certifications, and LinkedIn courses
- Opportunities to grow as a People Manager, technical specialist, Solution Architect, or Project or Delivery Manager
About the team
52.650
employees
Benefits and perks
- Flexible schedule
- Work from home
- Health insurance
- Training budget
- Seguro médico, dental y de visión
- Planes 401(k) y de jubilación
- Programa de asistencia al empleado
- Reembolso de matrícula y programas de desarrollo profesional
- Programas de bienestar y herramientas de mentoring