Padmi

Site Reliability Engineer Sre Lead Bengaluru (India)

BangalorePosted 1 month ago
Software engineeringSeniorFull Time; Regular
Apply at MANEVA CONSULTING

Opens the source posting on shine.com

Source description

About the role

View original

Greetings from Maneva! Job Description Job Title - Site Reliability Engineer (SRE) Lead Experience - 7 - 10 Years Location - PAN India Notice - Immediate Joiner Requirements: A Senior SRE drives enterprise-wide reliability, observability, automation, and operational excellence across large-scale hybrid settings. This level requires architectural judgement, leadership in high-severity incidents, and the ability to mature SRE practices. operational functions performed by Site Reliability Engineering teams to ensure availability, reliability, performance, and resilience of applications and infrastructure. Key Responsibilities Reliability Engineering & Service Governance Define SLIs/SLOs/SLAs and govern error budgets. Lead architecture reviews for reliability and resilience. Drive SRE adoption and reduce toil. Maintain and improve service availability, resilience, and SLO/SLI governance. Apply SRE principles (risk management, error budget tracking, elimination of toil). Evaluate application readiness and architecture for reliability standards. Observability & AIOps Architect observability platforms (AppDynamics, Datadog, Prometheus, Grafana, Splunk, Dynatrace, ELK). Implement logging, metrics, and tracing. Improve visibility, reduce noise, and optimize MTTD/MTTI/MTTR. Incident, Problem & Change Management Lead major incident response and RCA. Execute capacity, security, and change control processes. Improve operational processes (capacity, security, change). Automation & Platform Engineering Build automation for deployments, monitoring, remediation. Design CI/CD pipelines and IaC. Drive AIOps adoption. Leadership & Collaboration Mentor teams and lead reliability initiatives. Influence roadmaps and enforce reliability standards. Required Technical Expertise Solid hands-on experience in Unix/Linux, Shell scripting, Python/Java/Go/NodeJS. Deep expertise in monitoring & observability tools Distributed systems knowledge. Incident leadership and RCA. Proven CI/CD pipeline and IaC experience. Dashboard/KPI/metric creation and tracking. Experience Requirements 5+ years in SRE/DevOps/Production Engineering. Experience with mission-critical, large-scale systems. Ability to work with architects and leadership. Experience operating large-scale, distributed systems using SRE practices. If you are excited to grab this chance, please apply directly or share your CV at and . Greetings from Maneva! Job Description Job Title - Site Reliability Engineer (SRE) Lead Experience - 7 - 10 Years Location - PAN India Notice - Immediate Joiner Requirements: A Senior SRE drives enterprise-wide reliability, observability, automation, and operational excellence across large-scale hybrid settings. This level requires architectural judgement, leadership in high-severity incidents, and the ability to mature SRE practices. operational functions performed by Site Reliability Engineering teams to ensure availability, reliability, performance, and resilience of applications and infrastructure. Key Responsibilities Reliability Engineering & Service Governance Define SLIs/SLOs/SLAs and govern error budgets. Lead architecture reviews for reliability and resilience. Drive SRE adoption and reduce toil. Maintain and improve service availability, resilience, and SLO/SLI governance. Apply SRE principles (risk management, error budget tracking, elimination of toil). Evaluate application readiness and architecture for reliability standards. Observability & AIOps Architect observability platforms (AppDynamics, Datadog, Prometheus, Grafana, Splunk, Dynatrace, ELK). Implement logging, metrics, and tracing. Improve visibility, reduce noise, and optimize MTTD/MTTI/MTTR. Incident, Problem & Change Management Lead major incident response and RCA. Execute capacity, security, and change control processes. Improve operational processes (capacity, security, change). Automation & Platform Engineering Build automation for deployments, monitoring, remediation. Design CI/CD pipelines and IaC. Drive AIOps adoption. Leadership & Collaboration Mentor teams and lead reliability initiatives. Influence roadmaps and enforce reliability standards. Required Technical Expertise Solid hands-on experience in Unix/Linux, Shell scripting, Python/Java/Go/NodeJS. Deep expertise in monitoring & observability tools Distributed systems knowledge. Incident leadership and RCA. Proven CI/CD pipeline and IaC experience. Dashboard/KPI/metric creation and tracking. Experience Requirements 5+ years in SRE/DevOps/Production Engineering. Experience with mission-critical, large-scale systems. Ability to work with architects and leadership. Experience operating large-scale, distributed systems using SRE practices. If you are e

One address, no account. We’ll tell you when matching roles go live.

More at MANEVA CONSULTING

Related open roles

View all roles
Site Reliability Engineer Sre Lead Bengaluru (India) at MANEVA CONSULTING · Padmi