Source description
About the role
Job Description SRE | DevOps & Cloud Operations (5-6 Years Experience) Overview We are looking for a Site Reliability Engineer (SRE) with 5-6 years of experience in DevOps, Cloud Operations, ITOM, and Production Support. The ideal candidate will be responsible for ensuring platform reliability, application availability, monitoring, incident management, automation, and operational excellence across Azure and Kubernetes environments. Key Responsibilities - Manage Incident, Problem, Change, and Release Management processes. - Perform Root Cause Analysis (RCA) and drive preventive actions. - Administer and support Azure cloud infrastructure and Kubernetes environments. - Implement and maintain monitoring solutions using Grafana, Prometheus, Loki, and Nagios. - Monitor application availability, performance, and infrastructure health. - Manage ITSM/ITOM processes using ServiceNow or similar tools. - Automate infrastructure and operational tasks using Terraform and Ansible. - Support CI/CD deployments and operational automation. - Drive cloud cost optimization initiatives (FinOps). - Participate in on-call support and major incident management. Required Skills - 5-6 years of experience in SRE, DevOps, Cloud Operations, or Production Support. - Strong knowledge of: - Incident, Problem & Change Management - ITIL Processes - ITSM Tools (ServiceNow or equivalent) - IT Operations Management (ITOM) - Hands-on experience with: - Azure Cloud Administration & Operations - Kubernetes Administration - Grafana, Prometheus, Loki, Nagios - Terraform & Ansible - Linux Administration & Scripting - Application & Infrastructure Monitoring - Root Cause Analysis (RCA) - FinOps & Cloud Cost Optimization - CI/CD using GitHub Actions - Robust troubleshooting and production support experience. Preferred Skills - Experience with GitHub Actions, Azure DevOps, Jenkins, or GitLab CI/CD. - Python or Shell scripting for automation. - Docker and containerized application deployments. - DevSecOps practices and infrastructure governance. - Azure, Kubernetes, Terraform, or ITIL certifications. - Experience working in Agile/Scrum environments. Support Requirements - Based on project requirements, the candidate may be required to work in the CET (Central European Time) time zone - Service support may extend to weekend based on operational needs. - On-call support required during Indian public holidays to support operational activities. .
More at BMW TechWorks
Related open roles
SRE - DevOps & Cloud Operations - SWF-BNE
Mumbai
AI for DevOps initiative Steering Lead (Tamil Nadu)
India
Embedded C++ Engineer - Lights
Mumbai
Web App Full Stack Developer GenAI & Knowledge Systems (Bengaluru)
Bangalore
HWVE_ Java Backend Developer
Chennai
Agentic AI engineer for BMW innovation Solutions
Bangalore