Padmi

SRE

MumbaiPosted 2 months ago
Software engineeringSeniorFull Time; Regular
Apply at IIT JOBS INC

Opens the source posting on shine.com

Source description

About the role

View original

Site Reliability Engineer to own cloud infrastructure, automation, and system reliability in a hybrid Mumbai or Bangalore role. Responsibilities Manage high-severity incidents and implement preventive measures to reduce failure rates. Define and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Design and maintain distributed systems with a focus on consistency, fault tolerance, and scalability. Build and maintain continuous integration and deployment pipelines using modern tooling. Automate manual operational tasks using strong coding skills in Python, Go, or Java. Required Skills 6+ years of experience in Site Reliability Engineering or similar infrastructure roles. Deep hands-on expertise with GCP, including compute, storage, networking, and managed services. Proficiency with Azure is a plus. Strong experience with Terraform for infrastructure as code. Hands-on experience with configuration management tools such as GitHub, ArgoCD, Ansible, Chef, or Puppet. Experience with container orchestration using Kubernetes, Docker, or Podman. Practical experience with observability tools like Prometheus and Grafana. Understanding of distributed system design patterns and deployment strategies. Preferred Skills Immediate joiner available within 30 days. Hybrid work capability in Mumbai or Bangalore. Site Reliability Engineer to own cloud infrastructure, automation, and system reliability in a hybrid Mumbai or Bangalore role. Responsibilities Manage high-severity incidents and implement preventive measures to reduce failure rates. Define and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Design and maintain distributed systems with a focus on consistency, fault tolerance, and scalability. Build and maintain continuous integration and deployment pipelines using modern tooling. Automate manual operational tasks using strong coding skills in Python, Go, or Java. Required Skills 6+ years of experience in Site Reliability Engineering or similar infrastructure roles. Deep hands-on expertise with GCP, including compute, storage, networking, and managed services. Proficiency with Azure is a plus. Strong experience with Terraform for infrastructure as code. Hands-on experience with configuration management tools such as GitHub, ArgoCD, Ansible, Chef, or Puppet. Experience with container orchestration using Kubernetes, Docker, or Podman. Practical experience with observability tools like Prometheus and Grafana. Understanding of distributed system design patterns and deployment strategies. Preferred Skills Immediate joiner available within 30 days. Hybrid work capability in Mumbai or Bangalore.

One address, no account. We’ll tell you when matching roles go live.

More at IIT JOBS INC

Related open roles

View all roles