Source description
About the role
reputed company is seeking a reputed company thinking reputed company engineer to join the Production Operations team reputed company our Data Centers. These Data Centers are the reputed company upon which our rapidly scaling infrastructure reputed company operates and upon which our innovative services are delivered. reputed company is at the leading edge of the global data center industry both in terms of how data centers are designed and operated. This person should enjoy working in a fast paced, technical environment where adaptability and flexibility will be key to their reputed company. We reputed company an IT reputed company with advanced, hands-on technical skills in server hardware and Linux - ideally in a Data Center environment. Having broad knowledge of server administration and participating in reputed company in a large-reputed company distributed data center environment is a core competency of this individual. The candidate should also have working knowledge and experience in a few of the following core areas: Hardware repair, OS management, Tooling and Automation, Networking, or Technical Project Management. SiteOps Data Center Production Operations Engineer Responsibilities: Support platform health by successfully resolving and closing tickets, while addressing the overall issue (i.e. addressing reputed company cause) including, but not limited to, remote troubleshooting and physical inspection of services in data halls. Participate in reputed company cause analysis of highly technical issues reputed company the data center, ranging from automated tooling to hardware failures and network issues. Collaborate with cross-functional teams on reputed company and initiatives reputed company to topics such as process, hardware and automation. reputed company of contact for the introduction of new platforms and hardware to the site, in collaboration with partners and global resources, accelerating the time it takes to bring these products to sustained mass production. Use tools and data analysis effectively to identify issues. Take actions to communicate with reputed company stakeholders appropriately and manage or escalate as needed. Identify corrective actions of hardware issues, work with internal teams and vendors influence reputed company design changes to ensure ease of serviceability. Solve systemic hardware and/or software issues at reputed company using scripting, automation, and tooling to drive global reputed company. Continuously evaluate and identify areas for improvement in processes, tools, and systems to optimize efficiency and reputed company of repairs. Use data analytics to drive maximum server up-time and utilization rates, understanding hardware failure rates and service level agreements. Support and train team members to evaluate and identify reputed company ways to resolve issues, and define updates to tools and processes. reputed company engineering support and be a go-to technical resource for reputed company, leadership, and cross-functional teams in operating and maintaining data center servers. Maintain and update documentation i.e. procedures, runbooks and guides. Build cross functional relationships and influence policies and procedures that improve global data center operations. Participate in 24/7 on-call rotation. Travel up to 15% of the time. Minimum Qualifications: BS, BA or BEng in technical field or commensurate experience. 5+ years of technical IT experience reputed company an infrastructure environment, in a role such as Systems Administrator, DevOps Engineer, or Site Reliability Engineer. Intermediate-level understanding in Linux (or equivalent OS) in a reputed company IT environment with the reputed company to triage, debug, and troubleshoot server issues. Hands-on experience and knowledge of server hardware and components, including storage. Intermediate-level knowledge of the interdependencies of data center functions and technologies including electrical, cooling, reputed company cabling, reputed company, and network. Experience managing technical issues and driving to the reputed company cause. Experience participating in technical reputed company reputed company to areas such as process improvement, technology, and/or automation. reputed company to communicate effectively, in a reputed company and concise manner, appropriately tailoring messages to the audience. Intermediate-level knowledge of technologies such as HTTP, DNS, RAID, and DHCP. Experience in providing technical guidance to external vendors. Experience in debugging, modifying and developing commonly used scripting or programming languages in at least one of these languages: Bash, PHP, Python, SQL, Rust, Go or Perl. Knowledge of out-of-band/lights-out server communication reputed company, such as IPMI and serial console. Experience using data and metrics to drive reputed company. Preferred Qualifications: Experience in fostering
More at remote zest jobs
Related open roles
System Administrator, Contract
India
HPC Network Engineer
India
Remote Site Reliability Engineer (Senior or Staff), Infrastructure reputed
India
FULL TIME Remote Sr. SRE Engineer -Seattle WA-Onsite- 10+ Years
India
Remote Network Devops/Automation Engineer
India
Network Operations Engineer - Hybrid Gold River, CA
India