Source description
About the role
As a member of the Application SRE Team, your role will involve supporting critical components of foundational technologies for real-time protection, RBI, and SSPM services. You will be part of a team dedicated to improving availability, latency, performance, efficiency, change management, monitoring, emergency response, and capacity planning of the engineering stacks. If you enjoy solving complex problems and developing cloud services at scale, we are interested in speaking with you. What's in it for you: - Join a high-caliber engineering team in the dynamic field of cloud tools and infrastructure management. - Work on hybrid cloud environments (Google Cloud, On-prem cloud) utilizing cutting-edge tools like Spinnaker, Kubernetes, Docker, and more. - Tackle challenging problems to enhance your technical and analytical skills. - Make significant contributions to our market-leading product support, impacting our rapidly-growing global customer base. Your responsibilities will include: - Collaborating with development teams and product managers to design and build highly available, performant, and secure features. - Implementing innovative methods to monitor and report application and infrastructure health. - Acquiring in-depth knowledge of the application stack. - Enhancing the performance of microservices and addressing scaling/performance issues. - Engaging in capacity management and planning. - Operating effectively in a fast-paced and dynamic environment. - Participating in 24X7 on-call rotations with development teams. - Debugging, optimizing code, and automating routine tasks. - Improving system efficiencies through capacity planning, configuration management, performance tuning, monitoring, and root cause analysis. Required skills and experience: - Troubleshooting experience with Unix/Linux for 5+ years. - Background in managing large-scale web operations. - Proficiency in one or more of the following languages: C, C++, Java, Python, Go, Perl, or Ruby. - Familiarity with algorithms, data structures, complexity analysis, and software design. - Hands-on experience with private or public cloud services in a highly available and scalable production environment. - Knowledge of continuous integration and deployment automation tools such as Jenkins, Ansible, etc. - Understanding of distributed systems is advantageous. - Previous experience collaborating with geographically-distributed teams. - Strong interpersonal communication skills and the ability to work effectively in a diverse, team-oriented environment with SREs, developers, Product Managers, etc. - Leadership experience in delivering complex software features and solutions through cross-functional collaboration. Education: - BSCS or equivalent required, MSCS or equivalent strongly preferred. As a member of the Application SRE Team, your role will involve supporting critical components of foundational technologies for real-time protection, RBI, and SSPM services. You will be part of a team dedicated to improving availability, latency, performance, efficiency, change management, monitoring, emergency response, and capacity planning of the engineering stacks. If you enjoy solving complex problems and developing cloud services at scale, we are interested in speaking with you. What's in it for you: - Join a high-caliber engineering team in the dynamic field of cloud tools and infrastructure management. - Work on hybrid cloud environments (Google Cloud, On-prem cloud) utilizing cutting-edge tools like Spinnaker, Kubernetes, Docker, and more. - Tackle challenging problems to enhance your technical and analytical skills. - Make significant contributions to our market-leading product support, impacting our rapidly-growing global customer base. Your responsibilities will include: - Collaborating with development teams and product managers to design and build highly available, performant, and secure features. - Implementing innovative methods to monitor and report application and infrastructure health. - Acquiring in-depth knowledge of the application stack. - Enhancing the performance of microservices and addressing scaling/performance issues. - Engaging in capacity management and planning. - Operating effectively in a fast-paced and dynamic environment. - Participating in 24X7 on-call rotations with development teams. - Debugging, optimizing code, and automating routine tasks. - Improving system efficiencies through capacity planning, configuration management, performance tuning, monitoring, and root cause analysis. Required skills and experience: - Troubleshooting experience with Unix/Linux for 5+ years. - Background in managing large-scale web operations. - Proficiency in one or more of the following languages: C, C++, Java, Python, Go, Perl, or Ruby. - Familiarity with algorithms, data structures, complexity analysis, and software design. - Hands-on experience with private or public cloud services in a highly available and scalable production environ
More at Netskope