Source description
About the role
Site Reliability Engineer to own cloud infrastructure, automation, and system reliability in a hybrid Mumbai or Bangalore role. Responsibilities Manage high-severity incidents and implement preventive measures to reduce failure rates. Define and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Design and maintain distributed systems with a focus on consistency, fault tolerance, and scalability. Build and maintain continuous integration and deployment pipelines using modern tooling. Automate manual operational tasks using strong coding skills in Python, Go, or Java. Required Skills 6+ years of experience in Site Reliability Engineering or similar infrastructure roles. Deep hands-on expertise with GCP, including compute, storage, networking, and managed services. Proficiency with Azure is a plus. Strong experience with Terraform for infrastructure as code. Hands-on experience with configuration management tools such as GitHub, ArgoCD, Ansible, Chef, or Puppet. Experience with container orchestration using Kubernetes, Docker, or Podman. Practical experience with observability tools like Prometheus and Grafana. Understanding of distributed system design patterns and deployment strategies. Preferred Skills Immediate joiner available within 30 days. Hybrid work capability in Mumbai or Bangalore. Site Reliability Engineer to own cloud infrastructure, automation, and system reliability in a hybrid Mumbai or Bangalore role. Responsibilities Manage high-severity incidents and implement preventive measures to reduce failure rates. Define and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Design and maintain distributed systems with a focus on consistency, fault tolerance, and scalability. Build and maintain continuous integration and deployment pipelines using modern tooling. Automate manual operational tasks using strong coding skills in Python, Go, or Java. Required Skills 6+ years of experience in Site Reliability Engineering or similar infrastructure roles. Deep hands-on expertise with GCP, including compute, storage, networking, and managed services. Proficiency with Azure is a plus. Strong experience with Terraform for infrastructure as code. Hands-on experience with configuration management tools such as GitHub, ArgoCD, Ansible, Chef, or Puppet. Experience with container orchestration using Kubernetes, Docker, or Podman. Practical experience with observability tools like Prometheus and Grafana. Understanding of distributed system design patterns and deployment strategies. Preferred Skills Immediate joiner available within 30 days. Hybrid work capability in Mumbai or Bangalore.
More at IIT JOBS INC