Source description
About the role
Hiring for: A leading global IT Services & Consulting firm (CMM level 5 Company) Job Title: Site Reliability Engineer (SRE) With GCP Experience: 5+ Years Location: Mumbai Notice Period: Immediate Joiners / Serving Notice Period Preferred Job Summary We are seeking a highly skilled Site Reliability Engineer (SRE) with strong expertise in Kubernetes (GKE), ELK Stack, Apache Kafka, Redis Sentinel, and GCP . The ideal candidate will have hands-on experience managing large-scale production environments, automating infrastructure, and ensuring high availability, scalability, and reliability of critical platforms. Key Responsibilities Design, manage, and optimize Google Kubernetes Engine (GKE) clusters. Manage and administer ELK Stack (Elasticsearch, Logstash, Kibana) environments. Onboard new applications and log sources into ELK. Configure and manage RBAC, security roles, dashboards, and alerting in Kibana. Deploy and manage Redis Sentinel/Redis Clusters for high availability. Administer Apache Kafka clusters , including topic creation, replication management, and partition balancing. Build automation solutions using Python or Java . Perform Proof of Concepts (POCs) for multi-tier application automation. Manage cloud infrastructure on Google Cloud Platform (GCP) . Implement Infrastructure as Code (IaC) using Terraform and Ansible . Support CI/CD pipelines using ArgoCD, Jenkins, or GitLab CI/CD . Ensure platform reliability, performance optimization, monitoring, and operational excellence. Mandatory Skills Site Reliability Engineering (SRE) Kubernetes (GKE) Docker Elasticsearch Logstash Kibana ELK Stack Administration Apache Kafka Redis Sentinel / Redis Cluster Google Cloud Platform (GCP) Terraform Ansible Python or Java ArgoCD / Jenkins / GitLab CI/CD Interested candidates can share their updated resume to: sugan.g@taggd.in
More at Talent Hired-the Job Store