Padmi

DevOps - Site Reliability Engineer - AWS

Delhi NCRPosted 3 months ago
Infrastructure And DatabasesMid-levelFull Time; Regular
Apply at Innovaccer Analytics Private Limited

Opens the source posting on shine.com

Source description

About the role

View original

As a DevOps/SRE Engineer, you will be responsible for building a CICD stack in collaboration with the Dev and QA/Automation teams to drive the organization to a new level of continuous delivery and deployment. Security is a top priority, and you will work closely with the CISO and Dev teams to make security a first-class citizen. Your tasks will include developing S-CICD (Secure CICD), enabling various security tool chains, and providing vulnerability reports to developers through automation. Observability is critical for the scalability of our systems, and you will lead the effort to enhance observability spanning across logs, metrics, mesh, tracing, etc. Collaborating with Dev and QA teams, you will drive initiatives to increase the adoption of DevOps practices and tool chains. Your strong analytical skills will be crucial in understanding production system metrics, optimizing system utilization, and driving cost efficiency. During peak seasons, you will be responsible for auto-scaling the platform up or down accordingly. Ensuring that the platform adheres to security guidelines established by the CISO, such as securing against DDoS attacks and implementing WAF, vulnerability and patch management, and installing required security agents. Implementing least privilege-based RBAC for various production services and tool chains will be a key aspect of your role. You will also be involved in building and executing the Disaster Recovery plan and participating in Incident Response scenarios. Qualification Required: - 3+ years of experience as a DevOps/SRE Engineer - Solid experience with at least one cloud platform - AWS, Azure, GCP with automation focus. Certification is advantageous. - Hands-on experience with Kubernetes and Linux - Proficiency in scripting languages like Python - Experience in building scalable CICD architectures and solutions - Preferred experience in building observability stacks from logs, metrics, traces, service mesh, data observability - Good documentation skills for structuring documents for consumption by various dev teams - Cloud Security knowledge is highly preferred - Hands-on experience with technologies like Kafka, Postgre, Snowflake, etc. Additional Details: - Multi Cloud experience with AWS, Azure, GCP - Distributed Compute experience with Kubernetes (EKS/AKS), Containerization - Persistence stores experience with Postgres, MongoDB - Data Warehousing experience with Snowflake, Data Bricks - Messaging experience with Kafka - CICD experience with Jenkins, ArgoCD, GitOps - Observability experience with Elasticsearch, Prometheus, Jaeger, NewRelic, etc. As a DevOps/SRE Engineer, you will be responsible for building a CICD stack in collaboration with the Dev and QA/Automation teams to drive the organization to a new level of continuous delivery and deployment. Security is a top priority, and you will work closely with the CISO and Dev teams to make security a first-class citizen. Your tasks will include developing S-CICD (Secure CICD), enabling various security tool chains, and providing vulnerability reports to developers through automation. Observability is critical for the scalability of our systems, and you will lead the effort to enhance observability spanning across logs, metrics, mesh, tracing, etc. Collaborating with Dev and QA teams, you will drive initiatives to increase the adoption of DevOps practices and tool chains. Your strong analytical skills will be crucial in understanding production system metrics, optimizing system utilization, and driving cost efficiency. During peak seasons, you will be responsible for auto-scaling the platform up or down accordingly. Ensuring that the platform adheres to security guidelines established by the CISO, such as securing against DDoS attacks and implementing WAF, vulnerability and patch management, and installing required security agents. Implementing least privilege-based RBAC for various production services and tool chains will be a key aspect of your role. You will also be involved in building and executing the Disaster Recovery plan and participating in Incident Response scenarios. Qualification Required: - 3+ years of experience as a DevOps/SRE Engineer - Solid experience with at least one cloud platform - AWS, Azure, GCP with automation focus. Certification is advantageous. - Hands-on experience with Kubernetes and Linux - Proficiency in scripting languages like Python - Experience in building scalable CICD architectures and solutions - Preferred experience in building observability stacks from logs, metrics, traces, service mesh, data observability - Good documentation skills for structuring documents for consumption by various dev teams - Cloud Security knowledge is highly preferred - Hands-on experience with technologies like Kafka, Postgre, Snowflake, etc. Additional Details: - Multi Cloud experience with AWS, Azure, GCP - Distributed Compute experience with Kubernetes (EKS/AKS), Containerization - Persistence

One address, no account. We’ll tell you when matching roles go live.

More at Innovaccer Analytics Private Limited

Related open roles

View all roles