Padmi

Senior TL SRE

HyderabadPosted 2 months ago
Software engineeringSeniorFull Time; Regular
Apply at VERTAFORE INC

Opens the source posting on shine.com

Source description

About the role

View original

As a Senior Technical Lead Site Reliability Engineer at our company, you will be responsible for owning the reliability, scalability, performance, and operational integrity of critical production services. Your main tasks will include: - Owning production services end to end, ensuring reliability, availability, scalability, performance, and operational health. - Defining and managing SLIs and SLOs to guide delivery decisions. - Influencing service and system design to enhance fault tolerance, observability, and operational sustainability. - Debugging complex production issues across application code, services, and infrastructure using software engineering practices. - Performing root cause analysis using logs, metrics, traces, and code-level investigation. - Building automation and self-healing mechanisms to prevent repeat failures. - Executing production changes with safety, automation, and observability. - Designing and operating production observability aligned to service health and customer impact. - Leading and participating in incident response for high-severity events. - Collaborating with engineering, product, architecture, and operations teams. - Operating with autonomy and sound judgment in reliability decisions. To qualify for this role, you should have: - 8 to 12 years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership. - Proven operation at a senior engineering scope with accountability for reliability outcomes. - Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies. - Practical experience applying SRE principles such as SLIs, SLOs, and error budgets. - Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code, and hybrid environments. - Strong knowledge of Linux and Windows systems, application platforms, and relational databases. - Bachelors or masters degree in computer science or equivalent experience. - Participation in an on-call rotation and flexibility with working hours as required. Please note that Vertafore conducts preemployment background screenings and the selected candidate must be legally authorized to work in the United States. Vertafore strongly supports equal employment opportunity for all applicants regardless of various characteristics protected by law. As a Senior Technical Lead Site Reliability Engineer at our company, you will be responsible for owning the reliability, scalability, performance, and operational integrity of critical production services. Your main tasks will include: - Owning production services end to end, ensuring reliability, availability, scalability, performance, and operational health. - Defining and managing SLIs and SLOs to guide delivery decisions. - Influencing service and system design to enhance fault tolerance, observability, and operational sustainability. - Debugging complex production issues across application code, services, and infrastructure using software engineering practices. - Performing root cause analysis using logs, metrics, traces, and code-level investigation. - Building automation and self-healing mechanisms to prevent repeat failures. - Executing production changes with safety, automation, and observability. - Designing and operating production observability aligned to service health and customer impact. - Leading and participating in incident response for high-severity events. - Collaborating with engineering, product, architecture, and operations teams. - Operating with autonomy and sound judgment in reliability decisions. To qualify for this role, you should have: - 8 to 12 years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership. - Proven operation at a senior engineering scope with accountability for reliability outcomes. - Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies. - Practical experience applying SRE principles such as SLIs, SLOs, and error budgets. - Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code, and hybrid environments. - Strong knowledge of Linux and Windows systems, application platforms, and relational databases. - Bachelors or masters degree in computer science or equivalent experience. - Participation in an on-call rotation and flexibility with working hours as required. Please note that Vertafore conducts preemployment background screenings and the selected candidate must be legally authorized to work in the United States. Vertafore strongly supports equal employment opportunity for all applicants regardless of various characteristics protected by law.

One address, no account. We’ll tell you when matching roles go live.

More at VERTAFORE INC

Related open roles

View all roles
Senior TL SRE at VERTAFORE INC · Padmi