Source description
About the role
As a Senior Site Reliability Engineer Real Time Engineering at LSEG, you will be responsible for supporting the Real Time application space, ensuring operational stability, and contributing to cloud migration initiatives. Your role will involve automating operational tasks, building observability capabilities, and collaborating with global teams in a fast-paced environment handling real-time data. Key Responsibilities: - Lead incident recovery efforts by analyzing complex situations and coordinating effective recovery actions. - Apply structured problem-solving and critical thinking to identify root causes in distributed systems. - Produce clear incident reports and participate in post-incident reviews for reliability improvements. - Maintain production stability by assessing deployment risks and identifying reliability concerns using operational data. - Automate operational processes and enhance observability, alerting, and troubleshooting with AI-assisted tools. - Develop a strong understanding of application dataflows, dependencies, and networking topologies. - Collaborate with development teams on backlog prioritization, project intake, and operational acceptance testing. Technical & Professional Qualifications: - Experience in UNIX/Linux administration, scripting, and automation. - Hands-on experience supporting cloud-native applications, preferably on Azure; AWS experience is advantageous. - Understanding of networking fundamentals such as TCP/IP, UDP traffic, HTTP, and DNS. - Proficiency with Kubernetes, Docker, and container-based platforms. - Ability to troubleshoot distributed systems using logical reasoning and data-driven analysis. - Familiarity with observability tools; experience with DataDog and BigPanda is highly desirable. - Knowledge of Git or similar version control systems. - Bachelor's degree in computer science or related field, or equivalent practical experience. - Relevant industry experience; customer-facing operational support experience is a plus. Joining LSEG means being part of a global financial markets infrastructure and data provider committed to driving financial stability and empowering economies. The company values diversity, individuality, and innovation while fostering a culture of sustainability and inclusivity. LSEG offers tailored benefits, support, and opportunities for personal and professional growth to its employees. Please note that LSEG is an equal opportunities employer and does not discriminate based on any factors protected under applicable law. The company accommodates religious practices, beliefs, as well as mental health or physical disability needs. As a Senior Site Reliability Engineer Real Time Engineering at LSEG, you will be responsible for supporting the Real Time application space, ensuring operational stability, and contributing to cloud migration initiatives. Your role will involve automating operational tasks, building observability capabilities, and collaborating with global teams in a fast-paced environment handling real-time data. Key Responsibilities: - Lead incident recovery efforts by analyzing complex situations and coordinating effective recovery actions. - Apply structured problem-solving and critical thinking to identify root causes in distributed systems. - Produce clear incident reports and participate in post-incident reviews for reliability improvements. - Maintain production stability by assessing deployment risks and identifying reliability concerns using operational data. - Automate operational processes and enhance observability, alerting, and troubleshooting with AI-assisted tools. - Develop a strong understanding of application dataflows, dependencies, and networking topologies. - Collaborate with development teams on backlog prioritization, project intake, and operational acceptance testing. Technical & Professional Qualifications: - Experience in UNIX/Linux administration, scripting, and automation. - Hands-on experience supporting cloud-native applications, preferably on Azure; AWS experience is advantageous. - Understanding of networking fundamentals such as TCP/IP, UDP traffic, HTTP, and DNS. - Proficiency with Kubernetes, Docker, and container-based platforms. - Ability to troubleshoot distributed systems using logical reasoning and data-driven analysis. - Familiarity with observability tools; experience with DataDog and BigPanda is highly desirable. - Knowledge of Git or similar version control systems. - Bachelor's degree in computer science or related field, or equivalent practical experience. - Relevant industry experience; customer-facing operational support experience is a plus. Joining LSEG means being part of a global financial markets infrastructure and data provider committed to driving financial stability and empowering economies. The company values diversity, individuality, and innovation while fostering a culture of sustainability and inclusivity. LSEG offers tailored benefits, sup
More at LSEG