Apptad

Senior Site Reliability Engineer / Platform Engineer - AWS EKS (Hybrid)

Apptad

Hyderabad, Telangana, India (Hybrid) · Full Time

Be the first to apply

Experience
6+ yrs
Salary
—
Openings
1
Posted
1 day ago
Work mode
Hybrid
Education
Any graduate
Eligibility
Candidates holding any graduate degree are eligible to apply for this position.
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About the Role

This position is responsible for ensuring the reliability, security, automation, and scalability of the AWS Elastic Kubernetes Service (EKS) platform. Acting as a platform owner, the Senior Site Reliability Engineer / Platform Engineer supports application teams by providing standardized infrastructure, GitOps workflows, policy enforcement, and monitoring capabilities. The role demands advanced hands-on experience in Kubernetes, AWS cloud services, security tools, and SRE practices.

Work Location & Mode

  • Available work locations: Bangalore, Hyderabad, Pune
  • Work mode: Hybrid

Key Responsibilities

  • Design, deploy, and maintain production-grade Kubernetes environments on AWS EKS.
  • Automate EKS deployments with EKS Blueprints and manage lifecycle of EKS add-ons.
  • Plan and carry out Kubernetes and EKS version upgrades with minimal downtime.
  • Develop and operate autoscaling solutions using Karpenter for both stateless and stateful workloads.
  • Optimize cloud resource usage for cost, performance, and availability by dynamic instance provisioning.
  • Implement and manage Istio service mesh including sidecar and ambient mesh deployments.
  • Configure traffic policies such as mutual TLS, retries, circuit breaking, and timeouts within Istio.
  • Apply Kubernetes policy enforcement and admission control via Kyverno and OPA/Gatekeeper.
  • Operate Falco for runtime security monitoring and threat detection.
  • Integrate security controls within GitOps pipelines.
  • Create reusable Terraform modules for AWS infrastructure components such as VPCs, EKS clusters, and Transit Gateways.
  • Use Terragrunt to handle multi-account and multi-region infrastructure setups.
  • Manage Argo CD for platform, security policies, and add-on component delivery.
  • Define Git-based environment promotion workflows and access management strategies.
  • Implement Prometheus-based metrics monitoring and alerting systems.
  • Take part in incident response actions, perform root cause analyses, and drive reliability improvements.
  • Automate operations to reduce manual effort and enable self-service tooling.
  • Handle remediation of Wiz security alerts for AWS infrastructure and Kubernetes clusters.
  • Collaborate with security teams to establish preventive measures and guardrails.

Required Experience and Skills

  • Minimum 6 years in Site Reliability Engineering, Platform Engineering, or Cloud Infrastructure roles.
  • Strong expertise managing Kubernetes platforms and AWS EKS environments.
  • Proficient with Terraform, Terragrunt, and EKS Blueprints implementations.
  • Experience deploying Karpenter autoscaling for cloud workloads.
  • In-depth knowledge of Istio service mesh and its traffic management policies.
  • Practical experience with Kyverno, OPA/Gatekeeper for admission control and governance.
  • Using Falco for runtime security threat detection and forensic investigations.
  • Skilled in Prometheus for observability and alerting solutions.
  • Hands-on experience configuring GitOps workflows using Argo CD.
  • Solid understanding of AWS fundamentals including VPC, IAM, EC2, ALB/NLB, EBS, and EFS.
  • Proven ability to analyze and resolve Wiz security findings in AWS environments.

Preferred Qualifications

  • Experience with Ambient Mesh deployment models.
  • Familiarity with Site Reliability Engineering indicators such as SLOs, SLIs, and error budgets.
  • Exposure to operating within large-scale or regulated industries.

Eligibility

Applicants must hold any graduate degree to be eligible for this role.

Minimum education

Bachelor's Degree

Tools & software

AWS Amazon Web Services AWS required
🤖
Online · instant AI help
Broxer