Source description
About the role
Job Summary:Experienced Systems Administrator with a strong foundation in Linux, infrastructure management, and incident response, skilled in monitoring, troubleshooting, and maintaining reliable systems across virtualized and cloud-based environments.Job ResponsibilitiesManage and optimize Linux systems with focus on performance, reliability, and troubleshooting.Handle network issues including latency, packet drops, and connectivityWork on cloud platforms (AWS/GCP/Azure) for deployment and scalingDeploy and manage applications using Docker and Kubernetes (cluster troubleshooting & scaling)Build and maintain monitoring systems using Prometheus, Grafana, and ELKCreate dashboards, alerts, and PromQL queriesAutomate tasks using Python/Bash scriptingManage CI/CD pipelines (Jenkins/GitLab CI)Handle P1/P2 incidents, lead bridges, and perform RCAKey SkillsStrong Linux fundamentals.Good understanding of networking (TCP/IP, DNS, HTTP/HTTPS, load balancing)Hands-on experience with Docker & Kubernetes (must-have)Experience with cloud platforms (AWS/GCP/Azure)Knowledge of monitoring tools (Prometheus, Grafana, ELK)Proficiency in Python or Bash scriptingExperience in CI/CD tools (Jenkins/GitLab CI)Strong incident management and troubleshooting skills.Good to Have:Exposure to Terraform or AnsibleQualifications:Bachelors degree in Computer Science, Engineering (BE/B.Tech), MCA, or M.Sc (IT). Job Summary:Experienced Systems Administrator with a strong foundation in Linux, infrastructure management, and incident response, skilled in monitoring, troubleshooting, and maintaining reliable systems across virtualized and cloud-based environments.Job ResponsibilitiesManage and optimize Linux systems with focus on performance, reliability, and troubleshooting.Handle network issues including latency, packet drops, and connectivityWork on cloud platforms (AWS/GCP/Azure) for deployment and scalingDeploy and manage applications using Docker and Kubernetes (cluster troubleshooting & scaling)Build and maintain monitoring systems using Prometheus, Grafana, and ELKCreate dashboards, alerts, and PromQL queriesAutomate tasks using Python/Bash scriptingManage CI/CD pipelines (Jenkins/GitLab CI)Handle P1/P2 incidents, lead bridges, and perform RCAKey SkillsStrong Linux fundamentals.Good understanding of networking (TCP/IP, DNS, HTTP/HTTPS, load balancing)Hands-on experience with Docker & Kubernetes (must-have)Experience with cloud platforms (AWS/GCP/Azure)Knowledge of monitoring tools (Prometheus, Grafana, ELK)Proficiency in Python or Bash scriptingExperience in CI/CD tools (Jenkins/GitLab CI)Strong incident management and troubleshooting skills.Good to Have:Exposure to Terraform or AnsibleQualifications:Bachelors degree in Computer Science, Engineering (BE/B.Tech), MCA, or M.Sc (IT).
More at Crest Data