Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
EPAM provides enterprise software products, open source solutions, and technology accelerators through its SolutionsHub ecosystem.
задачи
Design, implement, and maintain production Kubernetes clusters on EKS and AKS across multiple environments and regions;
Configure and operate GitOps workflows with ArgoCD for declarative application delivery and cluster configuration;
Develop and manage Infrastructure-as-Code with Terraform, Helm, and Helmfile;
Ensure observability and reliability using Prometheus, Grafana, and CloudWatch;
Implement monitoring and alerting strategies;
Apply security best practices, including network policies, RBAC, pod security controls, and secret management;
Operate and optimize Kubernetes add-ons such as Argo Workflows, Argo Rollouts, cert-manager, CSI drivers, and cluster autoscalers;
Diagnose and resolve distributed systems issues across compute, networking, and storage layers;
Collaborate with development teams to improve deployment models, scalability, and resource utilization;
Document architectures, operational workflows, and incident handling standards.
требования
7+ Years of professional engineering experience with modern cloud-native architectures;
Strong expertise in AWS and Azure, including EKS, AKS, VPC/VNet, IAM, EBS/EFS, S3, Load Balancers, and Route53;
Deep understanding of Kubernetes architecture, networking, storage, and security principles with production-scale management experience;
Proven experience with Infrastructure-as-Code workflows using Terraform, including module development and state handling, and Helm/Helmfile;
Practical knowledge of GitOps tools such as ArgoCD or Flux and declarative infrastructure deployment;
Experience implementing security policies, secrets management, and compliance automation;
Advanced skill in observability tooling such as Prometheus and Grafana, and log management through CloudWatch or ELK/EFK;
Scripting proficiency in Bash, Python, or Go for automation and system tooling;
Strong knowledge of containerization and image lifecycle using Docker, ECR, or ACR;
Clear communication skills and ability to document complex systems effectively;
Nice to have: Familiarity with policy-as-code frameworks such as Kyverno and OPA/Gatekeeper, experience with AWS PrivateLink, VPC peering, and hybrid/multi-cloud networking, understanding of service mesh patterns, ingress controllers, and DNS management, exposure to performance tuning and capacity planning for Kubernetes environments.