Source description
About the role
About the Role We are seeking a skilled and proactive Site Reliability Engineer (SRE). This role involves close collaboration with developers and QA teams, ensuring seamless transitions and ongoing reliability of applications. Responsibilities - Daily monitoring, incident response, and performance tuning - Automation and optimization to reduce manual effort - Ensuring platform reliability - Supporting efforts to manage and contain increasing cloud costs - Managing the rapidly growing data estate in Azure - Willingness to work in on-call rotations or provide after-hours support if needed Requirements: - Should have total 8+ years of experience and 4+ years of relevant experience. - Knowledge of SDLC process like requirement gathering, design, implementation (Coding), testing, deployment, and maintenance. - Proficient in requirement gathering and documentation - Excellent communication skills verbal and written. - Azure Platform Operations & Monitoring - Hands-on experience with Azure Monitor, Log Analytics, and Application Insights - Familiarity with Azure Data Factory, Synapse, and Databricks - Working knowledge of Azure Storage, Key Vault, and Role-Based Access Control (RBAC) - Monitoring & Incident Management - Proficiency in setting up alerts, dashboards, and automated responses - Experience with uptime/health monitoring and SLA enforcement - Ability to triage platform issues and coordinate with engineering teams - Proficient with CI/CD pipelines and infrastructure-as-code (e.g., Azure DevOps, Terraform) - Understanding of data pipeline dependencies and scheduling (especially ADF triggers) - Comfortable with log review, root cause analysis, and performance tuning - Strong problem-solving and communication skills - Willingness to work in on-call rotations or provide after-hours support if needed - Familiarity with ITIL or operational runbooks is a valuable to have. We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our or please contact .Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams . Please read our . About the Role We are seeking a skilled and proactive Site Reliability Engineer (SRE). This role involves close collaboration with developers and QA teams, ensuring seamless transitions and ongoing reliability of applications. Responsibilities - Daily monitoring, incident response, and performance tuning - Automation and optimization to reduce manual effort - Ensuring platform reliability - Supporting efforts to manage and contain increasing cloud costs - Managing the rapidly growing data estate in Azure - Willingness to work in on-call rotations or provide after-hours support if needed Requirements: - Should have total 8+ years of experience and 4+ years of relevant experience. - Knowledge of SDLC process like requirement gathering, design, implementation (Coding), testing, deployment, and maintenance. - Proficient in requirement gathering and documentation - Excellent communication skills verbal and written. - Azure Platform Operations & Monitoring - Hands-on experience with Azure Monitor, Log Analytics, and Application Insights - Familiarity with Azure Data Factory, Synapse, and Databricks - Working knowledge of Azure Storage, Key Vault, and Role-Based Access Control (RBAC) - Monitoring & Incident Management - Proficiency in setting up alerts, dashboards, and automated responses - Experience with uptime/health monitoring and SLA enforcement - Ability to triage platform issues and coordinate with engineering teams - Proficient with CI/CD pipelines and infrastructure-as-code (e.g., Azure DevOps, Terraform) - Understanding of data pipeline dependencies and scheduling (especially ADF triggers) - Comfortable with log review, root cause analysis, and performance tuning - Strong problem-solving and communication skills - Willingness to work in on-call rotations or provide after-hours support if needed - Familiarity with ITIL or operational runbooks is a valuable to have. We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our or please contact .Criminals may pose as recruiters asking for money or personal information. We never request money or banking details from job applicants. Learn more about spotting and avoiding scams . Please read our .