Padmi

Senior Site Reliability Engineer (SRE) (Cloud & Kubernetes)

HyderabadPosted 2 months ago
Software engineeringSeniorFull Time; Regular
Apply at Adecco India

Opens the source posting on shine.com

Source description

About the role

View original

Job Title: Senior Site Reliability Engineer (SRE) Location: Hyderabad (Hybrid Work Model) Experience: 4 - 8 Years Salary: Negotiable Education : Bachelors Degree / Masters Degree in Computer Science, Information Technology, or related field. Employment Type: Full Time Job Description We are looking for an experienced Senior Site Reliability Engineer (SRE) to join our Infrastructure & Engineering team. The ideal candidate should have strong experience in SRE practices, Java application deployments, Kubernetes, GCP, and observability tools. The role involves ensuring the reliability, scalability, availability, and performance of critical production systems while collaborating closely with development and infrastructure teams. Roles & Responsibilities Maintain high availability, scalability, and reliability of production environments Collaborate with developers to improve deployment processes and operational excellence Manage and improve CI/CD pipelines for secure and reliable deployments Implement and manage Infrastructure as Code (IaC) using Terraform and related tools Work on containerization technologies like Docker and Kubernetes Configure monitoring, logging, dashboards, and alerting systems Lead incident response, troubleshooting, root cause analysis, and reliability improvements Improve observability through metrics, tracing, dashboards, and alerts Support Java application deployments in cloud-native environments Participate in on-call support and production issue management Conduct capacity planning, reliability reviews, and automation initiatives Drive SRE best practices including SLIs/SLOs, error budgets, and toil reduction Mandatory Skills Prior experience working in a Site Reliability Engineer (SRE) role Strong Java application deployment experience Hands-on experience in Java application development Strong experience with Docker and Kubernetes Hands-on experience with Google Cloud Platform (GCP) Experience in logging, monitoring, alerting, dashboards, and observability tools Strong understanding of CI/CD pipelines Linux and shell scripting experience Knowledge of networking concepts such as TCP/IP, DNS, HTTP, TLS/SSL Good to Have Skills Terraform, Helm, Jenkins, GitHub Actions, ArgoCD Service Mesh technologies like Istio Microservices architecture experience Python scripting Experience with SLIs/SLOs and reliability engineering practices Chaos testing and capacity planning E-commerce domain experience Interested candidates kindly share your CV and below details to usha.sundar@adecco.com Present CTC (Fixed + VP) - Expected CTC - No. of years experience - Notice Period - Offer-in hand - Reason of Change - Present Location - Job Title: Senior Site Reliability Engineer (SRE) Location: Hyderabad (Hybrid Work Model) Experience: 4 - 8 Years Salary: Negotiable Education : Bachelors Degree / Masters Degree in Computer Science, Information Technology, or related field. Employment Type: Full Time Job Description We are looking for an experienced Senior Site Reliability Engineer (SRE) to join our Infrastructure & Engineering team. The ideal candidate should have strong experience in SRE practices, Java application deployments, Kubernetes, GCP, and observability tools. The role involves ensuring the reliability, scalability, availability, and performance of critical production systems while collaborating closely with development and infrastructure teams. Roles & Responsibilities Maintain high availability, scalability, and reliability of production environments Collaborate with developers to improve deployment processes and operational excellence Manage and improve CI/CD pipelines for secure and reliable deployments Implement and manage Infrastructure as Code (IaC) using Terraform and related tools Work on containerization technologies like Docker and Kubernetes Configure monitoring, logging, dashboards, and alerting systems Lead incident response, troubleshooting, root cause analysis, and reliability improvements Improve observability through metrics, tracing, dashboards, and alerts Support Java application deployments in cloud-native environments Participate in on-call support and production issue management Conduct capacity planning, reliability reviews, and automation initiatives Drive SRE best practices including SLIs/SLOs, error budgets, and toil reduction Mandatory Skills Prior experience working in a Site Reliability Engineer (SRE) role Strong Java application deployment experience Hands-on experience in Java application development Strong experience with Docker and Kubernetes Hands-on experience with Google Cloud Platform (GCP) Experience in logging, monitoring, alerting, dashboards, and observability tools Strong understanding of CI/CD pipelines Linux and shell scripting experience Knowledge of networking concepts such as TCP/IP, DNS, HTTP, TLS/SSL Good to Have Skills Terraform, Helm, Jenkins, GitHub Actions, ArgoCD Service Mesh technologies like Istio Microserv

One address, no account. We’ll tell you when matching roles go live.

More at Adecco India

Related open roles

View all roles