Padmi

Track Lead - Kubernetes, Terraform

Delhi NCRPosted 1 month ago
Infrastructure And DatabasesMid-levelFull Time; Regular
Apply at HCLTech

Opens the source posting on shine.com

Source description

About the role

View original

Noida, Uttar Pradesh Job Summary The Linux L3 Engineer is responsible for advanced administration, troubleshooting, and optimization of Linux systems across on-premises and cloud platforms. The role includes owning complex incidents, driving automation, ensuring system reliability, and supporting hybrid infrastructure services integrated with storage, cloud, virtualization, and enterprise tooling Key Responsibilities Perform advanced administration of Linux servers (RHEL, CentOS, Ubuntu) in enterprise environments. Handle L3-level troubleshooting for OS, CPU, memory, disk, kernel, and performance issues. Manage OS patching, upgrades, kernel tuning, and image lifecycle. Perform server provisioning, decommissioning, and configuration management. Own and resolve P1/P2 incidents related to Linux systems. Perform root cause analysis (RCA) and implement preventive actions. Support change, release, and configuration management processes. Proactively monitor system performance and ensure availability. Analyze CPU, memory, IO, filesystem usage and optimize system performance. Implement proactive measures and automation for self-healing systems. Support Linux workloads on AWS and OCI, including: o VPC, networking, load balancers, storage (ZFS, EBS-like services) o Monitoring and alerting Manage Linux-based workloads on IaaS and PaaS platforms, including container hosts. Build and support infrastructure provisioning using Terraform (IaaS). Develop automation scripts (Shell/Python) for operational efficiency. Support Kubernetes-based platforms (SRE/PRE scope). Manage and troubleshoot container host environments. Support Linux integration with NAS solutions such as Nasuni Filers. Monitor filer health, cache utilization ( Skill Requirements Operating Systems: RHEL, CentOS, Ubuntu Core Linux Skills: LVM, filesystem management, systemd, networking Kernel tuning, performance troubleshooting Scripting & Automation: Shell scripting, Python Cloud Platforms: AWS, OCI Infrastructure as Code: Terraform Containers: Kubernetes, Docker Monitoring Tools: Nagios, Prometheus, CloudWatch, OEM tools Storage & NAS: Nasuni, NFS, CIFS Virtualization: VMware (basic integration knowledge) Other Requirements Operating Systems: RHEL, CentOS, Ubuntu Core Linux Skills: LVM, filesystem management, systemd, networking Kernel tuning, performance troubleshooting Scripting & Automation: Shell scripting, Python Cloud Platforms: AWS, OCI Infrastructure as Code: Terraform Containers: Kubernetes, Docker Monitoring Tools: Nagios, Prometheus, CloudWatch, OEM tools Storage & NAS: Nasuni, NFS, CIFS Virtualization: VMware (basic integration knowledge) Strong experience in: Experience in hybrid cloud environments (on-prem + cloud) Exposure to DevOps/SRE practices Knowledge of configuration management tools (Ansible, Puppet, Chef) Familiarity with enterprise backup solutions (Dell ABS, Data Domain) Understanding of networking concepts (DNS, TCP/IP, firewall basics) Experience in event correlation and monitoring optimization Hands-on with: VMware Aria Operations (vROps), Aria Automation, Log Insight Dell VxRail / PowerEdge infrastructure Storage (vSAN, NAS, Dell Unity, ZFS basic integration) Lifecycle Manager (LCM), patching & upgrades Monitoring & alerting tools (Datadog, SolarWinds, etc.) Basic scripting: PowerCLI / PowerShell #body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply- .

One address, no account. We’ll tell you when matching roles go live.

More at HCLTech

Related open roles

View all roles