Source description
About the role
As a Senior DevOps Engineer at BMC, you will be part of the Cloud Platform team responsible for architecting, deploying, maintaining, and supporting infrastructure platforms across public, private, and hybrid cloud environments. The core platform services owned by the team include Kubernetes, Infrastructure as Code, CI/CD platforms, and GitOps tooling for data center and cloud-native deployments. Your key responsibilities will include: - Supporting, maintaining, and optimizing private and public cloud deployments across AWS, Google Cloud, Oracle Cloud, and other IaaS platforms. - Maintaining and supporting platform deployment software such as Kubernetes, Rancher, OKE, GKE, and EKS. - Designing, developing, and maintaining automation using Terraform, Jenkins, ArgoCD, and Ansible. - Writing and maintaining automation and tooling using Python and/or Go. - Deploying and operating data center environments across bare metal, virtualized, and cloud-based infrastructure. - Providing architectural input for capacity planning, performance management, data center rollouts, and consolidation initiatives. - Designing and implementing AI-assisted automation for operational workflows including alert enrichment, runbooks, self-healing workflows, troubleshooting, and incident summarization. - Integrating AI capabilities into existing platform tooling, CI/CD pipelines, and ChatOps workflows (Slack/Teams). - Partnering with SRE, Ops, and Platform stakeholders to identify automation opportunities and drive operational improvements. - Ensuring all automation and AI solutions are secure, reliable, auditable, and production-ready. To ensure your success in this role, you should have the following qualifications and experience: - Strong hands-on experience with Terraform, Jenkins, Kubernetes, and ArgoCD. - 7-10 years of relevant industry experience. - Proven experience with production platforms and incident response. - Experience with Change and Incident Management tools and processes. - Prior experience in deploying and operating infrastructure in major cloud providers like AWS, Google Cloud, or Oracle Cloud. - Strong Linux fundamentals and system troubleshooting skills. - Proficiency in Python for building AI-assisted tooling, automation workflows, and integrations. - Hands-on experience using AI/LLM APIs for practical automation use cases. - Solid understanding of AI limitations, guardrails, and operational risks. - Ability to work independently with sound technical judgment. - Strong communication skills for collaboration with a global team. - Familiarity with Kubernetes management platforms, observability stacks, AIOps platforms, and event-driven automation systems. - Exposure to frameworks such as LangChain, LlamaIndex, or vector databases. - Relevant certifications like ITIL, AWS, GCP, OCI, VCP, etc. As a Senior DevOps Engineer at BMC, you will be part of the Cloud Platform team responsible for architecting, deploying, maintaining, and supporting infrastructure platforms across public, private, and hybrid cloud environments. The core platform services owned by the team include Kubernetes, Infrastructure as Code, CI/CD platforms, and GitOps tooling for data center and cloud-native deployments. Your key responsibilities will include: - Supporting, maintaining, and optimizing private and public cloud deployments across AWS, Google Cloud, Oracle Cloud, and other IaaS platforms. - Maintaining and supporting platform deployment software such as Kubernetes, Rancher, OKE, GKE, and EKS. - Designing, developing, and maintaining automation using Terraform, Jenkins, ArgoCD, and Ansible. - Writing and maintaining automation and tooling using Python and/or Go. - Deploying and operating data center environments across bare metal, virtualized, and cloud-based infrastructure. - Providing architectural input for capacity planning, performance management, data center rollouts, and consolidation initiatives. - Designing and implementing AI-assisted automation for operational workflows including alert enrichment, runbooks, self-healing workflows, troubleshooting, and incident summarization. - Integrating AI capabilities into existing platform tooling, CI/CD pipelines, and ChatOps workflows (Slack/Teams). - Partnering with SRE, Ops, and Platform stakeholders to identify automation opportunities and drive operational improvements. - Ensuring all automation and AI solutions are secure, reliable, auditable, and production-ready. To ensure your success in this role, you should have the following qualifications and experience: - Strong hands-on experience with Terraform, Jenkins, Kubernetes, and ArgoCD. - 7-10 years of relevant industry experience. - Proven experience with production platforms and incident response. - Experience with Change and Incident Management tools and processes. - Prior experience in deploying and operating infrastructure in major cloud providers like AWS, Google Cloud, or Oracle Cloud. - Strong Linux fundamentals and syste
More at BMC Software