Source description
About the role
Job Responsibilities
Provide thought leadership in support of associated hardware, and systems software planning and automated process improvements, environment preparation, and production support.
Ability to translate business requirements into technical implementation and operational management.
Passionate about building, observing and operating distributed systems at scale in production.
Understand the challenges and trade-offs to be made when building and deploying new systems to production.
Collaborate with a team to build, review and help design high quality software and systems.
Have experience in secure multi-tenant cloud development
Responsible for developing tools which provides health and operational insights of services, which drives business decisions.
Responsible for the development and maintenance of operations based on product and services. Map and align operations monitoring to Services Health to ensure end to end availability of services and applications.
Validated foundation in RESTful APIs, data structures, algorithms, software design and security.
Enthusiastic about making the many users of your product happier every day.
Expected to be a self-starter, innovative and always looking for new ways to contribute to the team.
Lead, mentor, and partner with team members and other groups. Monitoring of infra services and applications, including designing and implementing automated methods to selfheal the network and unified communications infrastructure Analyze and improve availability, efficiency, capacity, scalability and performance of services. Work with Service Owners and partners to agree metrics and performance through the delivery of proactive monitoring services. Developing and maintaining operations tools. Introduce improvements by leveraging automation and innovative approaches to Service Health and Management. Basic Qualifications Bachelor's Degree in any technical discipline. 5+ years of technical experience working in complex IT environments focused on automation tools 5+ years of experience in application development or production support. 2+ years of experience in infrastructure automation tools with good coding experience in one or more of python, Java or C#. Preferred Qualification
Passion for ensuring systems are highly available.
Auto-scaling, Infrastructure as code, automated monitoring and reporting
You have worked extensively with container deployment and orchestration technologies at scale with knowledge of the fundamentals to include service discovery, deployments, monitoring, scheduling, load balancing. Knowledge of Python, Java and PowerShell are preferred.
Experience with development and deployment in a hosted cloud environment, preferably Microsoft Azure.
Experience with building SaaS multi-tenant applications using OAuth or SAML authentication providers.
Experience working in ETL tools (Extract, Transform and load data set to structured format)
Experience working in TSDS tools (Time Series Database)
Experience defining the metrics and integrating the applications with analytics engines like ELK stack/ Splunk
Ability to develop, manage and communicate frameworks. Strong analytical and problem-solving skills to create new solutions, planning, scheduling, and organization skills. Familiarity & understanding of public and private cloud architecture. Strong understanding of networking, compute infrastructure, distributed storage and Internet protocols
The ability to learn new technologies quickly and provide mentorship.
Strong interpersonal skills, both verbal and written
Job type: Full Time/Contract Division: eTeam Workforce Limited Reference: 22-05549
More at eTeam