Padmi

Manager, Site Reliability Engineering - 6sense

BangalorePosted 2 months ago
Technology ManagementSeniorFull Time; Regular
Apply at OpenTalent

Opens the source posting on shine.com

Source description

About the role

View original

Key Responsibilities Lead and mentor a team of SREs, driving best practices in reliability, automation, and incident management.Architect and maintain highly available, scalable Kubernetes clusters on AWS, ensuring secure and efficient deployments.Design and implement CI/CD pipelines, monitoring, and alerting systems to support rapid, reliable releases.Collaborate with product, engineering, and security teams to define SLAs, SLOs, and capacity planning.Own postmortem processes, rootcause analysis, and continuous improvement initiatives.Requirements 5+ years of SRE or DevOps experience in a fastmoving SaaS environment.Deep expertise with Kubernetes, AWS services (EKS, EC2, RDS, CloudWatch), and container orchestration.Proven track record building CI/CD pipelines (GitHub Actions, ArgoCD, Jenkins) and observability stacks (Prometheus, Grafana, Loki).Strong leadership skills, with experience managing and scaling engineering teams.Excellent communication, problemsolving, and incident response abilities. Key Responsibilities Lead and mentor a team of SREs, driving best practices in reliability, automation, and incident management.Architect and maintain highly available, scalable Kubernetes clusters on AWS, ensuring secure and efficient deployments.Design and implement CI/CD pipelines, monitoring, and alerting systems to support rapid, reliable releases.Collaborate with product, engineering, and security teams to define SLAs, SLOs, and capacity planning.Own postmortem processes, rootcause analysis, and continuous improvement initiatives.Requirements 5+ years of SRE or DevOps experience in a fastmoving SaaS environment.Deep expertise with Kubernetes, AWS services (EKS, EC2, RDS, CloudWatch), and container orchestration.Proven track record building CI/CD pipelines (GitHub Actions, ArgoCD, Jenkins) and observability stacks (Prometheus, Grafana, Loki).Strong leadership skills, with experience managing and scaling engineering teams.Excellent communication, problemsolving, and incident response abilities.

One address, no account. We’ll tell you when matching roles go live.

More at OpenTalent

Related open roles

View all roles