Source description
About the role
As a Principal Site Reliability Engineer at UKG, you play a critical role in enhancing service delivery processes through software solutions. Your responsibilities include: - Engaging in and improving the lifecycle of services from conception to End of Life (EOL) - Defining and implementing standards and best practices related to system architecture, service delivery, metrics, and automation of operational tasks - Supporting services, product & engineering teams by providing common tooling and frameworks for increased availability and improved incident response - Improving system performance, application delivery, and efficiency through automation, process refinement, postmortem reviews, and configuration analysis - Collaborating closely with engineering professionals to deliver reliable services - Increasing operational efficiency and quality of services by treating operational challenges as a software engineering problem - Guiding junior team members and championing Site Reliability Engineering - Actively participating in incident response, including on-call responsibilities - Partnering with stakeholders to drive the best technical and business outcomes Qualifications required for this role include: - Engineering degree, or a related technical discipline, or equivalent work experience - Experience in coding in higher-level languages such as Python, JavaScript, C, or Java - Knowledge of cloud-based applications and containerization technologies - Understanding of best practices in metric generation, log aggregation pipelines, time-series databases, and distributed tracing - Fundamentals in Computer Science, Cloud Architecture, Security, or Network Design - Working experience with industry standards like Terraform, Ansible Additionally, you should have: - At least 7 years of hands-on experience in Engineering or Cloud - Minimum 5 years of experience with public cloud platforms (e.g., GCP, AWS, Azure) - Minimum 3 years of experience in configuration and maintenance of applications and/or systems infrastructure for large-scale customer-facing companies - Experience with distributed system design and architecture At UKG, you will have the opportunity to work with purpose and be part of a team that values inclusivity, engagement, and innovation. If you are passionate about driving flawless customer experiences and are eager to contribute to a leading global software company, then this role is the perfect fit for you. As a Principal Site Reliability Engineer at UKG, you play a critical role in enhancing service delivery processes through software solutions. Your responsibilities include: - Engaging in and improving the lifecycle of services from conception to End of Life (EOL) - Defining and implementing standards and best practices related to system architecture, service delivery, metrics, and automation of operational tasks - Supporting services, product & engineering teams by providing common tooling and frameworks for increased availability and improved incident response - Improving system performance, application delivery, and efficiency through automation, process refinement, postmortem reviews, and configuration analysis - Collaborating closely with engineering professionals to deliver reliable services - Increasing operational efficiency and quality of services by treating operational challenges as a software engineering problem - Guiding junior team members and championing Site Reliability Engineering - Actively participating in incident response, including on-call responsibilities - Partnering with stakeholders to drive the best technical and business outcomes Qualifications required for this role include: - Engineering degree, or a related technical discipline, or equivalent work experience - Experience in coding in higher-level languages such as Python, JavaScript, C, or Java - Knowledge of cloud-based applications and containerization technologies - Understanding of best practices in metric generation, log aggregation pipelines, time-series databases, and distributed tracing - Fundamentals in Computer Science, Cloud Architecture, Security, or Network Design - Working experience with industry standards like Terraform, Ansible Additionally, you should have: - At least 7 years of hands-on experience in Engineering or Cloud - Minimum 5 years of experience with public cloud platforms (e.g., GCP, AWS, Azure) - Minimum 3 years of experience in configuration and maintenance of applications and/or systems infrastructure for large-scale customer-facing companies - Experience with distributed system design and architecture At UKG, you will have the opportunity to work with purpose and be part of a team that values inclusivity, engagement, and innovation. If you are passionate about driving flawless customer experiences and are eager to contribute to a leading global software company, then this role is the perfect fit for you.
More at UKG