Padmi

Sr. Site Reliability Engineer

IndiaPosted 2 months ago
Infrastructure And DatabasesSeniorFull Time; Regular
Apply at Crest Data

Opens the source posting on shine.com

Source description

About the role

View original

Job Summary:Experienced Systems Administrator with a strong foundation in Linux, infrastructure management, and incident response, skilled in monitoring, troubleshooting, and maintaining reliable systems across virtualized and cloud-based environments.Job ResponsibilitiesManage and optimize Linux systems with focus on performance, reliability, and troubleshooting.Handle network issues including latency, packet drops, and connectivityWork on cloud platforms (AWS/GCP/Azure) for deployment and scalingDeploy and manage applications using Docker and Kubernetes (cluster troubleshooting & scaling)Build and maintain monitoring systems using Prometheus, Grafana, and ELKCreate dashboards, alerts, and PromQL queriesAutomate tasks using Python/Bash scriptingManage CI/CD pipelines (Jenkins/GitLab CI)Handle P1/P2 incidents, lead bridges, and perform RCAKey SkillsStrong Linux fundamentals.Good understanding of networking (TCP/IP, DNS, HTTP/HTTPS, load balancing)Hands-on experience with Docker & Kubernetes (must-have)Experience with cloud platforms (AWS/GCP/Azure)Knowledge of monitoring tools (Prometheus, Grafana, ELK)Proficiency in Python or Bash scriptingExperience in CI/CD tools (Jenkins/GitLab CI)Strong incident management and troubleshooting skills.Good to Have:Exposure to Terraform or AnsibleQualifications:Bachelors degree in Computer Science, Engineering (BE/B.Tech), MCA, or M.Sc (IT). Job Summary:Experienced Systems Administrator with a strong foundation in Linux, infrastructure management, and incident response, skilled in monitoring, troubleshooting, and maintaining reliable systems across virtualized and cloud-based environments.Job ResponsibilitiesManage and optimize Linux systems with focus on performance, reliability, and troubleshooting.Handle network issues including latency, packet drops, and connectivityWork on cloud platforms (AWS/GCP/Azure) for deployment and scalingDeploy and manage applications using Docker and Kubernetes (cluster troubleshooting & scaling)Build and maintain monitoring systems using Prometheus, Grafana, and ELKCreate dashboards, alerts, and PromQL queriesAutomate tasks using Python/Bash scriptingManage CI/CD pipelines (Jenkins/GitLab CI)Handle P1/P2 incidents, lead bridges, and perform RCAKey SkillsStrong Linux fundamentals.Good understanding of networking (TCP/IP, DNS, HTTP/HTTPS, load balancing)Hands-on experience with Docker & Kubernetes (must-have)Experience with cloud platforms (AWS/GCP/Azure)Knowledge of monitoring tools (Prometheus, Grafana, ELK)Proficiency in Python or Bash scriptingExperience in CI/CD tools (Jenkins/GitLab CI)Strong incident management and troubleshooting skills.Good to Have:Exposure to Terraform or AnsibleQualifications:Bachelors degree in Computer Science, Engineering (BE/B.Tech), MCA, or M.Sc (IT).

One address, no account. We’ll tell you when matching roles go live.

More at Crest Data

Related open roles

View all roles