About this role
Responsibilities Operate and maintain Kubernetes platforms based on AWS EKS and Rancher -managed RKE2 . Manage cluster health, upgrades, patching, and capacity lifecycle activities. Investigate incidents using logs and metrics to perform detailed Root Cause Analysis (RCA) . Support AWS infrastructure operations involving EC2 , VPC , IAM , and DNS . Enhance monitoring, alerting, and operational procedures for increased stability. Automate recurring tasks to ensure repeatable and efficient platform operations. Document troubleshooting steps, known issues, and standard operating procedures. Requirements 1+ years of experience in DevOps, cloud operations, or Linux engineering. Experience with Kubernetes and cloud-native platform concepts. Experience with AWS and Linux system administration. Knowledge of networking fundamentals including DNS , TLS , and load balancing. Experience with monitoring, logging, and troubleshooting tools. Ability to perform structured Root Cause Analysis and write technical documentation. You are fluent in English . Nice to Haves Experience with AWS EKS , Rancher , or RKE2 . Experience with GitHub , GitHub Actions , Helm , or GitOps . Familiarity with observability tools such as Prometheus , Grafana , or CloudWatch . Familiarity with container security and secrets management. Possession of Kubernetes or AWS certifications.