Source description
About the role
Work Experience - 10+ Work Location - Chennai, Mumbai and Gurugram. Work Mode - Hybrid. We are looking for an experienced Site Reliability Engineer / DevOps Engineer with strong expertise in Linux, Terraform, Azure Cloud, and Kubernetes (AKS) to join our team. The ideal candidate should have hands-on experience in infrastructure automation, Kubernetes cluster management, and CI/CD implementation . Key Responsibilities Manage and troubleshoot Linux-based systems and environments Develop and maintain Infrastructure as Code using Terraform , including writing Terraform modules from scratch Design and manage scalable infrastructure on Microsoft Azure Administer and maintain Kubernetes clusters , particularly Azure Kubernetes Service (AKS) Handle Kubernetes cluster lifecycle management (scaling, upgrades, troubleshooting, and maintenance) Build and maintain CI/CD pipelines , preferably using GitHub Actions Implement DevOps and SRE best practices to improve system reliability and automation Collaborate with development teams to streamline deployment and infrastructure processes Mandatory Skills Strong experience in Linux OS Hands-on experience with Terraform (module development) Experience working with Azure Cloud Expertise in Kubernetes cluster management (AKS) Experience with CI/CD tools (preferably GitHub Actions ) Good to Have Experience with monitoring tools such as ELK Stack, Prometheus, or Grafana Exposure to DevOps monitoring and automation practices Preferred Certifications Terraform Certification Microsoft Azure Certification (AZ-900 or higher) Kubernetes Certifications – CKA / CKAD / CKS Note: We are specifically looking for candidates with hands-on experience managing and maintaining Kubernetes clusters , not just deploying applications on AKS.
More at datum technologies group