Source description
About the role
As a Cluster Administrator, you will be responsible for managing and optimizing large-scale Elasticsearch and Splunk clusters across public cloud, Kubernetes, and private data centers. You will play a crucial role in evaluating cluster performance, applying security fixes, maintaining data pipeline components, and automating infrastructure to ensure efficient operations. Key Responsibilities: - Evaluate cluster performance to identify developing issues and plan for high-load events. - Apply security fixes, perform incremental upgrades, and extend monitoring and alert infrastructure. - Maintain data pipeline components, including serverless and server-based systems. - Automate infrastructure using declarative tools and scripting to achieve build-once/run-everywhere systems. - Participate in an on-call rotation to resolve escalated technical issues. Qualifications Required: - 5+ years of experience in cluster administration and management. - Expertise in managing Elasticsearch clusters. - Experience with Splunk cluster administration. - Proficiency in Puppet for configuration management. - Strong scripting skills in Python, Bash, Go, Ruby, Perl, or similar languages. - Experience with CI/CD tools including Jenkins, Pipelines, and Artifactory. - Ability to troubleshoot complex data flows within large, disparate systems. - Bachelor's Degree in Computer Science, IT, Engineering, or a related discipline. In addition to the above qualifications, preferred skills for this role include experience with Terraform, Packer, or CloudFormation for Infrastructure as Code, as well as knowledge of monitoring tools such as Prometheus and PagerDuty. As a Cluster Administrator, you will be responsible for managing and optimizing large-scale Elasticsearch and Splunk clusters across public cloud, Kubernetes, and private data centers. You will play a crucial role in evaluating cluster performance, applying security fixes, maintaining data pipeline components, and automating infrastructure to ensure efficient operations. Key Responsibilities: - Evaluate cluster performance to identify developing issues and plan for high-load events. - Apply security fixes, perform incremental upgrades, and extend monitoring and alert infrastructure. - Maintain data pipeline components, including serverless and server-based systems. - Automate infrastructure using declarative tools and scripting to achieve build-once/run-everywhere systems. - Participate in an on-call rotation to resolve escalated technical issues. Qualifications Required: - 5+ years of experience in cluster administration and management. - Expertise in managing Elasticsearch clusters. - Experience with Splunk cluster administration. - Proficiency in Puppet for configuration management. - Strong scripting skills in Python, Bash, Go, Ruby, Perl, or similar languages. - Experience with CI/CD tools including Jenkins, Pipelines, and Artifactory. - Ability to troubleshoot complex data flows within large, disparate systems. - Bachelor's Degree in Computer Science, IT, Engineering, or a related discipline. In addition to the above qualifications, preferred skills for this role include experience with Terraform, Packer, or CloudFormation for Infrastructure as Code, as well as knowledge of monitoring tools such as Prometheus and PagerDuty.
More at Persistent Systems