Padmi

Site Reliability Engineer (SRE) Kubernetes Platform

IndiaPosted 2 months ago
Software engineeringMid-levelFull Time; Regular
Apply at Cisco Systems

Opens the source posting on shine.com

Source description

About the role

View original

Role Overview: The Meraki SRE Platform Engineering team at Cisco is responsible for building and operating the infrastructure that supports Meraki's cloud services. The focus is on delivering reliable, scalable, and simple platforms that enable product teams to work efficiently while ensuring a secure operating environment. As a Site Reliability Engineer in this team, you will play a crucial role in supporting the development and operation of the Kubernetes-based platform across various environments. You will collaborate with senior engineers and technical leaders to enhance reliability, scalability, and compliance across the platform. This role is hands-on, involving contributions to key systems, infrastructure building for production operations, and implementation of solutions to enhance the platform's overall health and performance. Key Responsibilities: - Contribute to the design, implementation, and operation of Kubernetes platforms - Support day-to-day reliability and performance of platform services, including monitoring and alerting - Implement automation and tooling to enhance operational efficiency and reduce manual effort - Work with senior engineers to define and track SLIs, SLOs, and error budgets - Assist in maintaining compliance and security requirements, including support for audits and continuous monitoring - Contribute to infrastructure as code and CI/CD pipeline improvements - Collaborate with multi-functional teams (Security, Platform, Application teams) to resolve issues and deliver platform capabilities - Participate in on-call rotations supporting customer requests and paging alerts Qualifications Required: - 3-4 years of experience in SRE, DevOps, or platform engineering roles - Experience with Kubernetes in production environments - Familiarity with cloud platforms such as AWS, Azure, or similar - Solid understanding of Linux systems, networking, and containerization - Experience with Infrastructure as Code (e.g., Terraform) - Proficiency in scripting or programming (e.g., Python, Go) Additional Company Details: Cisco, as a technology company, is revolutionizing how data and infrastructure connect and protect organizations in the AI era and beyond. With a history of fearless innovation spanning 40 years, Cisco creates solutions that facilitate the collaboration between humans and technology in both physical and digital realms. The solutions offered provide customers with unmatched security, visibility, and insights across their entire digital footprint. The depth and breadth of Cisco's technology fuel experimentation and the creation of meaningful solutions. This, combined with a global network of experts, presents limitless opportunities for growth and development. Cisco operates as a team, emphasizing collaboration and empathy to drive significant global impact through innovative solutions. The influence of Cisco is pervasive, making the impact felt everywhere, and it all begins with you. Role Overview: The Meraki SRE Platform Engineering team at Cisco is responsible for building and operating the infrastructure that supports Meraki's cloud services. The focus is on delivering reliable, scalable, and simple platforms that enable product teams to work efficiently while ensuring a secure operating environment. As a Site Reliability Engineer in this team, you will play a crucial role in supporting the development and operation of the Kubernetes-based platform across various environments. You will collaborate with senior engineers and technical leaders to enhance reliability, scalability, and compliance across the platform. This role is hands-on, involving contributions to key systems, infrastructure building for production operations, and implementation of solutions to enhance the platform's overall health and performance. Key Responsibilities: - Contribute to the design, implementation, and operation of Kubernetes platforms - Support day-to-day reliability and performance of platform services, including monitoring and alerting - Implement automation and tooling to enhance operational efficiency and reduce manual effort - Work with senior engineers to define and track SLIs, SLOs, and error budgets - Assist in maintaining compliance and security requirements, including support for audits and continuous monitoring - Contribute to infrastructure as code and CI/CD pipeline improvements - Collaborate with multi-functional teams (Security, Platform, Application teams) to resolve issues and deliver platform capabilities - Participate in on-call rotations supporting customer requests and paging alerts Qualifications Required: - 3-4 years of experience in SRE, DevOps, or platform engineering roles - Experience with Kubernetes in production environments - Familiarity with cloud platforms such as AWS, Azure, or similar - Solid understanding of Linux systems, networking, and containerization - Experience with Infrastructure as Code (e.g., Terraform) - Proficiency in scripting or programmin

One address, no account. We’ll tell you when matching roles go live.

More at Cisco Systems

Related open roles

View all roles
Site Reliability Engineer (SRE) Kubernetes Platform at Cisco Systems · Padmi