Source description
About the role
Job Title: Site Reliability Engineer (SRE) / Cloud Reliability Engineer Experience Required: 5 - 7 Years Location: [Bangalore - On-site] Job Type: Full-Time Industry: [e.g., IT Services, SaaS, Enterprise Technology, Fintech] Job Description: We are looking for a hands-on Site Reliability Engineer (SRE) / Cloud Reliability Engineer to build, automate, and operate highly reliable, scalable, and secure cloud platforms. The ideal candidate will have strong expertise in cloud infrastructure, automation, Infrastructure as Code (IaC), containerization, and observability, with a focus on ensuring platform reliability, performance, and operational excellence. Key Skills: Critical Skills (Must-Have) Cloud Platforms: AWS (Primary) Containers: Docker (Critical), Kubernetes,Infrastructure as Code (IaC), Terraform (Critical). Programming & Automation: Python (Critical), Golang, Shell Scripting. AIOps & Observability: Monitoring and Alerting Automation using Datadog, Prometheus, Grafana, Incident Intelligence and Automation using AIOps practices. Core Competencies: Cloud Infrastructure Design and OperationsInfrastructure AutomationReliability EngineeringMonitoring, Alerting, and Incident ManagementCI/CD Support and Platform OperationsPerformance, Availability, and Scalability OptimizationGood to Have Kubernetes (EKS), HelmObservability tools: Prometheus, Grafana Preferred Profile Strong hands-on experience in cloud infrastructure and automation.Ability to troubleshoot complex production issues in distributed systems.Experience working in high-availability and mission-critical environments.Strong understanding of reliability, monitoring, and operational best practices.Excellent problem-solving and collaboration skills.Role: Site Reliability Engineer (SRE) / Cloud Reliability Engineer Primary Cloud Platform: AWS Focus Areas: Reliability Engineering, Cloud Operations, Infrastructure Automation, AIOps, and Observability. Key Responsibilities: Design, implement, and operate scalable, highly available cloud infrastructure.Automate infrastructure provisioning, configuration management, and deployment processes.Implement reliability engineering practices to improve system availability, performance, and resiliency.Build and maintain monitoring, alerting, and incident management solutions using AIOps practices.Support CI/CD pipelines and cloud platform operations to enable efficient software delivery.Troubleshoot infrastructure, application, and platform issues while ensuring minimal service disruption.Collaborate with development and operations teams to improve system reliability and operational efficiency.Drive automation initiatives to reduce manual effort and enhance operational scalabilityInterested share your resume to akshitha@ashratech.com/8688322632Pay: 411,335.33 - 1,629,489.25 per year Work Location: In person Job Title: Site Reliability Engineer (SRE) / Cloud Reliability Engineer Experience Required: 5 - 7 Years Location: [Bangalore - On-site] Job Type: Full-Time Industry: [e.g., IT Services, SaaS, Enterprise Technology, Fintech] Job Description: We are looking for a hands-on Site Reliability Engineer (SRE) / Cloud Reliability Engineer to build, automate, and operate highly reliable, scalable, and secure cloud platforms. The ideal candidate will have strong expertise in cloud infrastructure, automation, Infrastructure as Code (IaC), containerization, and observability, with a focus on ensuring platform reliability, performance, and operational excellence. Key Skills: Critical Skills (Must-Have) Cloud Platforms: AWS (Primary) Containers: Docker (Critical), Kubernetes,Infrastructure as Code (IaC), Terraform (Critical). Programming & Automation: Python (Critical), Golang, Shell Scripting. AIOps & Observability: Monitoring and Alerting Automation using Datadog, Prometheus, Grafana, Incident Intelligence and Automation using AIOps practices. Core Competencies: Cloud Infrastructure Design and OperationsInfrastructure AutomationReliability EngineeringMonitoring, Alerting, and Incident ManagementCI/CD Support and Platform OperationsPerformance, Availability, and Scalability OptimizationGood to Have Kubernetes (EKS), HelmObservability tools: Prometheus, Grafana Preferred Profile Strong hands-on experience in cloud infrastructure and automation.Ability to troubleshoot complex production issues in distributed systems.Experience working in high-availability and mission-critical environments.Strong understanding of reliability, monitoring, and operational best practices.Excellent problem-solving and collaboration skills.Role: Site Reliability Engineer (SRE) / Cloud Reliability Engineer Primary Cloud Platform: AWS Focus Areas: Reliability Engineering, Cloud Operations, Infrastructure Automation, AIOps, and Observability. Key Responsibilities: Design, implement, and operate scalable, highly available cloud infrastructure.Automate infrastructure provisioning, configuration management, and depl
More at Ashra Technology