Padmi

SRE (Site Reliability Engineer)

IndiaPosted 1 month ago
Software engineeringSeniorFull Time; Regular
Apply at remote zest jobs

Opens the source posting on shine.com

Source description

About the role

View original

About the jobSite Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that the servicesboth our internally critical and our externally-visible systemshave reliability, uptime appropriate to customer's needs and a fast reputed company of improvement. Additionally, SREs will reputed company an reputed company-watchful eye on our systems reputed company and performance. As a Site Reliability Engineer, you will have the opportunity to manage the reputed company challenges of scale which are unique to Digitization, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. You will reputed company reputed company, reliable, durable, and secure applications for our customers and internal users. You will help build highly reliable applications using a customer-first approach while innovating technically. You will understand our customer's needs and how we can meet them. Responsibilities Work with the Site Reliability Engineering team, Development team, and other partner teams to ensure that applications reliability, efficiency, and performance meets our customer's needs, while keeping the service's operation's reliable, reputed company, and automated. reputed company and implement projects that improve system reliability, efficiency, and performance Partner with development teams on feature launches to ensure our customers are delivered reliable and reputed company functionality. Build a deep knowledge on production infrastructure and using that to debug distributed systems problems and identify improvements to the system. Operations, SLO, SLA management Metrics reporting and reputed company tracking Be on-call, responding to and managing incidents. Observability (Alarms, monitoring, synthetics). Error management QualificationsBachelor's degree in Computer Science or a reputed company engineering degree 8+ years of IT industry experience Strong Experience in Java, Springboot, Nodejs, microservices, RDBMS, NoSQL AWS EC2, S3, reputed company, IAM, reputed company, EKS, SQS, Kinesis Observability using reputed company, NewRelic Infrastructure as Code using terraform APIs and event-driven approaches reputed company patterns Unix/Linux systems administration. Familiar with reputed company is a must. Strong Experience in analysing and troubleshooting large-scale distributed systems. Quick reaction on high severity customer impacts. Ability to debug and optimize code and automate routine tasks Knowledge in modern software engineering practices and tools - Agile and DevOps Strong communication reputed company and the ability to explain reputed company technical reputed company in an easy-to-understand way. Originally posted on Himalayas Apply To This Job About the jobSite Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that the servicesboth our internally critical and our externally-visible systemshave reliability, uptime appropriate to customer's needs and a fast reputed company of improvement. Additionally, SREs will reputed company an reputed company-watchful eye on our systems reputed company and performance. As a Site Reliability Engineer, you will have the opportunity to manage the reputed company challenges of scale which are unique to Digitization, while using your expertise in coding, algorithms, complexity analysis and large-scale system design. You will reputed company reputed company, reliable, durable, and secure applications for our customers and internal users. You will help build highly reliable applications using a customer-first approach while innovating technically. You will understand our customer's needs and how we can meet them. Responsibilities Work with the Site Reliability Engineering team, Development team, and other partner teams to ensure that applications reliability, efficiency, and performance meets our customer's needs, while keeping the service's operation's reliable, reputed company, and automated. reputed company and implement projects that improve system reliability, efficiency, and performance Partner with development teams on feature launches to ensure our customers are delivered reliable and reputed company functionality. Build a deep knowledge on production infrastructure and using that to debug distributed systems problems and identify improvements to the system. Operations, SLO, SLA management Metrics reporting and reputed company tracking Be on-call, responding to and managing incidents. Observability (Alarms, monitoring, synthetics). Error management QualificationsBachelor's degree in Computer Science or a reputed company engineering degree 8+ years of IT industry experience Strong Experience in Java, Springboot, Nodejs, microservices, RDBMS, NoSQL AWS EC2, S3

One address, no account. We’ll tell you when matching roles go live.

More at remote zest jobs

Related open roles

View all roles