Padmi

IN-Associate_SRE Engineer_Digital Engineering Transformation_Advisory

BangalorePosted 2 months ago
Software engineeringMid-levelFull Time; Regular
Apply at PwC Service Delivery Center

Opens the source posting on shine.com

Source description

About the role

View original

Line of Service Advisory Industry/Sector Not Applicable Specialism Operations Management Level Associate Job Description Summary At PwC, our people in business services and support focus on providing efficient and effective administrative support to enable smooth operations within the organisation. This includes managing schedules, coordinating meetings, and handling confidential information. Those in client administration at PwC will focus on managing and coordinating client relationships, prioritising smooth communication and efficient service delivery. You will utilise strong organisational skills and attention to detail to support the overall client experience. As an early member of PwCs Digital Engineering practice, youll work at the intersection of industrial software and reliability engineering helping global energy and industrial technology companies operationalize what good looks like for mission-critical systems. Youll work directly with engineering leaders to define SLI/SLO frameworks for platforms, build unified observability stacks across multi-cloud environments, and establish the incident management discipline their teams need to move from reactive ops to automated reliability. This is a foundational role youll shape how an entire practice thinks about SRE, not just execute a playbook. Responsibilities: Define and implement SLI/SLO frameworks for industrial software platforms APM, SCADA, edge and OT-connected systems Design and deploy observability pipelines using OpenTelemetry, Prometheus, Grafana, Loki, and Jaeger across AWS and Azure environments Establish incident management processes and lead blameless postmortems with client engineering teams Build and execute chaos engineering scenarios to surface reliability risks before they hit production Develop automated runbooks and self-healing patterns that reduce toil and on-call burden Build SRE capability within client organisations including documentation, knowledge transfer, and SRE culture coaching Mandatory skill sets: SLI/SLO design and error budget management at enterprise scale Observability toolchain: Prometheus, Grafana, Loki, Jaeger, OpenTelemetry ideally across multi-cloud (AWS + Azure) Incident management, on-call process design, and postmortem facilitation Automation and scripting: Python, Go, or Bash runbook development a plus GitOps and CI/CD SRE gate experience (ArgoCD, GitHub Actions, Flux) Preferred skill sets: Familiarity with industrial or OT/IT environments energy, manufacturing, or connected products Disclaimer : This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying. Line of Service Advisory Industry/Sector Not Applicable Specialism Operations Management Level Associate Job Description Summary At PwC, our people in business services and support focus on providing efficient and effective administrative support to enable smooth operations within the organisation. This includes managing schedules, coordinating meetings, and handling confidential information. Those in client administration at PwC will focus on managing and coordinating client relationships, prioritising smooth communication and efficient service delivery. You will utilise strong organisational skills and attention to detail to support the overall client experience. As an early member of PwCs Digital Engineering practice, youll work at the intersection of industrial software and reliability engineering helping global energy and industrial technology companies operationalize what good looks like for mission-critical systems. Youll work directly with engineering leaders to define SLI/SLO frameworks for platforms, build unified observability stacks across multi-cloud environments, and establish the incident management discipline their teams need to move from reactive ops to automated reliability. This is a foundational role youll shape how an entire practice thinks about SRE, not just execute a playbook. Responsibilities: Define and implement SLI/SLO frameworks for industrial software platforms APM, SCADA, edge and OT-connected systems Design and deploy observability pipelines using OpenTelemetry, Prometheus, Grafana, Loki, and Jaeger across AWS and Azure environments Establish incident management processes and lead blameless postmortems with client engineering teams Build and execute chaos engineering scenarios to surface reliability risks before they hit production Develop automated runbooks and self-healing patterns that reduce toil and on-call burden Build SRE capability within client organisations including documentation, knowledge transfer, and SRE culture coaching Mandatory skill sets: SLI/SLO design and error budget management at enterprise scale Observability toolchain: Prometheus, Grafana, Loki, Jaeger, OpenTelemetry ideally across multi-cloud

One address, no account. We’ll tell you when matching roles go live.

More at PwC Service Delivery Center

Related open roles

View all roles