Padmi

Senior Site Reliability Engineer Senior Devops Engineer - AWS

MumbaiPosted 2 months ago
Infrastructure And DatabasesSeniorFull Time; Regular
Apply at Prometteur Solutions Pvt. Ltd.

Opens the source posting on shine.com

Source description

About the role

View original

Role Overview: As a Senior DevOps Engineer / SRE at Prometteur, you will be responsible for taking end-to-end ownership of infrastructure reliability, scalability, security, and cost efficiency across IT services projects and product-based platforms. This role extends beyond pipeline management to include production stability, system observability, performance optimization, and cloud cost governance. You will be supporting applications developed using Node.js, JavaScript, React.js, React Native, and Python, deployed across monolithic, microservices, and serverless architectures. Key Responsibilities: - Own production infrastructure reliability, uptime, and performance across all environments - Architect, deploy, and maintain scalable, secure AWS infrastructure and dedicated servers - Design and manage containerized workloads using Docker and orchestration platforms (ECS / EKS / Kubernetes) - Establish and maintain monitoring, logging, and alerting systems for proactive incident detection - Lead incident response, root cause analysis (RCA), and post-mortems - Optimize AWS cloud costs through right-sizing, usage analysis, and architectural improvements - Support high-traffic and multi-tenant SaaS systems with a focus on isolation and scalability - Collaborate closely with backend, frontend, and mobile teams to improve reliability and deployment workflows Qualifications Required: - 6+ years of experience in DevOps / SRE / Platform Engineering - Strong hands-on expertise with AWS (EC2, VPC, IAM, RDS, S3, Lambda, CloudWatch, etc.) - Experience managing dedicated / bare-metal servers and Linux-based systems - Deep understanding of networking, DNS, SSL/TLS, load balancing, and firewalls - Strong experience with Docker - Hands-on experience with container orchestration platforms such as ECS / EKS / Kubernetes - Proven experience with monitoring and observability stacks such as Prometheus, Grafana, ELK / OpenSearch, Loki - Strong experience with GitHub and Git-based workflows - Hands-on expertise in building CI/CD pipelines using GitHub Actions - Strong scripting skills (Bash, Python preferred) - Production experience supporting applications built with Node.js, JavaScript, Python, React.js, and React Native - Strong understanding of Monolithic architectures, Microservices architectures, Serverless, and event-driven systems - Production experience with MongoDB and MySQL - Knowledge of replication, backup strategies, scaling, and performance tuning - Demonstrated experience in AWS cost optimization, including resource right-sizing, cost monitoring and forecasting, and eliminating unused or inefficient services - IAM and least-privilege access, Secrets management, Network isolation and encryption - Experience working with high-traffic platforms and/or multi-tenant SaaS architectures - Excellent communication skills with engineering and leadership teams Role Overview: As a Senior DevOps Engineer / SRE at Prometteur, you will be responsible for taking end-to-end ownership of infrastructure reliability, scalability, security, and cost efficiency across IT services projects and product-based platforms. This role extends beyond pipeline management to include production stability, system observability, performance optimization, and cloud cost governance. You will be supporting applications developed using Node.js, JavaScript, React.js, React Native, and Python, deployed across monolithic, microservices, and serverless architectures. Key Responsibilities: - Own production infrastructure reliability, uptime, and performance across all environments - Architect, deploy, and maintain scalable, secure AWS infrastructure and dedicated servers - Design and manage containerized workloads using Docker and orchestration platforms (ECS / EKS / Kubernetes) - Establish and maintain monitoring, logging, and alerting systems for proactive incident detection - Lead incident response, root cause analysis (RCA), and post-mortems - Optimize AWS cloud costs through right-sizing, usage analysis, and architectural improvements - Support high-traffic and multi-tenant SaaS systems with a focus on isolation and scalability - Collaborate closely with backend, frontend, and mobile teams to improve reliability and deployment workflows Qualifications Required: - 6+ years of experience in DevOps / SRE / Platform Engineering - Strong hands-on expertise with AWS (EC2, VPC, IAM, RDS, S3, Lambda, CloudWatch, etc.) - Experience managing dedicated / bare-metal servers and Linux-based systems - Deep understanding of networking, DNS, SSL/TLS, load balancing, and firewalls - Strong experience with Docker - Hands-on experience with container orchestration platforms such as ECS / EKS / Kubernetes - Proven experience with monitoring and observability stacks such as Prometheus, Grafana, ELK / OpenSearch, Loki - Strong experience with GitHub and Git-based workflows - Hands-on expertise in building CI/CD pipelines using GitHub Actions - Strong scripting skills (Bash

One address, no account. We’ll tell you when matching roles go live.

More at Prometteur Solutions Pvt. Ltd.

Related open roles

View all roles