Source description
About the role
-Senior OpenShift Kubernetes Administrator to help lead our cloud-native platform transformation. -Opportunity to engineer, automate, and operate enterprise-scale OpenShift environments that power mission-critical applications. -You will work with cutting-edge technologies, solve complex platform challenges, and influence the future direction of our Kubernetes strategy. -As a senior member of our Kubernetes Platform Engineering team, you will be responsible for the reliability, scalability, security, and automation of our OpenShift ecosystem. Key Responsibilities • Administer, maintain, and optimize enterprise OpenShift Kubernetes platforms deployed to our High-Performance Compute infrastructure. • Manage large-scale OpenShift clusters using: o OpenShift Console o OpenShift CLI (oc) o Advanced Cluster Manager (ACM) o Infrastructure automation tools • Design and implement automated deployment, configuration, and lifecycle management solutions. • Develop Infrastructure-as-Code and automation frameworks using Ansible and scripting. • Troubleshoot complex Kubernetes, container, networking, and platform issues. • Collaborate with developers to improve application reliability, observability, scalability, and performance. • Implement enterprise security controls and platform hardening standards. • Support high availability, disaster recovery, and multi-cluster architectures. • Drive platform modernization initiatives and operational excellence through automation. • Participate in architecture reviews and help define the future state of enterprise container platforms. Required Qualifications • 5+ years administering Kubernetes platforms in production environments. • Deep expertise with Red Hat OpenShift. • Strong knowledge of Kubernetes internals including: o Scheduling o Networking o Storage o Ingress and load balancing o Operators o Cluster lifecycle management • Experience with Ansible and automation frameworks. • Experience with container technologies and image creation. • Proficiency in shell scripting and at least one programming language such as Python or Go. • Strong Linux administration skills, preferably Red Hat Enterprise Linux. • Experience troubleshooting enterprise-scale platform issues. • Excellent communication and technical documentation skills. • Experience deploying and administering OpenShift Virtualization (KubeVirt) to support both containerized and virtual machine workloads. • Experience implementing and supporting enterprise monitoring, logging, and alerting solutions using Grafana, Prometheus, Loki, AlertManager, and Thanos. • Demonstrated security-first mindset with expertise in platform hardening, identity and access management, vulnerability remediation, encryption, and regulatory compliance. • Experience with Git Action pipeline deployments for OpenShift, bash scripting using oc for controls, deployments, and working closely with application teams on deployments, new namespaces, affinities, and sizing. • Experience with cross-datacenter, high availability failover, and load balancing (F5 & haproxy) between multiple datacenters and K8s clusters Preferred Qualifications • Red Hat OpenShift certification. • Experience with GitOps, CI/CD pipelines, and Infrastructure-as-Code. • Experience supporting high-availability Kubernetes environments. • Familiarity with GPFS (IBM Spectrum Scale) or other distributed storage technologies. • Knowledge of enterprise networking concepts and Kubernetes networking architectures. • Experience supporting mission-critical applications with strict uptime requirements. • Candidates with experience in any of the following will stand out: • OpenShift deployed to IBM LinuxONE and IBM Z environments • s390x Linux architecture • Work on one of the industry's largest OpenShift deployments. • Influence platform strategy and architecture decisions. • Solve complex scalability and reliability challenges. • Partner with elite engineers and architects. • Build technology that directly impacts healthcare outcomes for millions of people.