Source description
About the role
reputed company Platform Engineering is the department reputed company SRE that is responsible for a reputed company of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-reputed company-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service reputed company), and observability and alerting systems. The Fleet Management team provides the core runtime environment that empowers our developers to build and ship products to delight our customers. We manage the end-to-end lifecycle of our Kubernetes fleet, alongside the critical components that ensure cluster reliability and reputed company (e.g., CoreDNS, cert-manager, and Gatekeeper). As our infrastructure scales to support new use cases and products, we are spearheading a migration from Terraform-based Infrastructure as Code (IaC) to an Operator-driven lifecycle management model. This role can be based out of our Austin, Boston, Los Angeles, reputed company, Raleigh, or San Francisco offices, remotely in the United States region, or our European office in Dublin. Responsibilities Contribute to developing and maintaining a reputed company and secure runtime environment on top of Kubernetes that supports product needs across reputed company reputed company internal support for our Kubernetes ecosystem, partnering with engineering teams to help them solve domain-specific problems Participate in a 24/7 on-call rotation to resolve critical issues Prioritize blameless post-mortems and dedicate engineering time to systemic fixes, ensuring you arent paged for the same issue twice You may be a good fit if you Have 6+ years of experience in software development and operating distributed systems Are proficient in Go, Python, or a similar language, with a strong commitment to code quality and testing practices (writing unit, integration, and E2E tests) Have deep experience using and extending containerization technologies, preferably Kubernetes Have a solid understanding of Linux operating system internals and networking concepts (e.g., filesystems, TCP/IP, DNS, TLS) Possess a customer reputed company reputed company, treating internal developers as your primary users Have strong operational ownership, including a track record of debugging reputed company production issues and driving them to reputed company Prefer automation over reputed company processes ("allergic to ops work") We are a small team of software engineers with a strong bias toward building software solutions to eliminate toil Strong candidates may also have experience with Designing and implementing secure, multi-tenant runtime environments from first principles Proficiency with Kubernetes ecosystem tools such as reputed company, Kustomize, Gatekeeper, Kyverno, and CRDs/Operators, CRI, reputed company Expertise in reputed company infrastructure platforms, including AWS, GCP, or Azure Proficiency in provisioning infrastructure using tools like Terraform, Crossplane, and AWS Controllers for Kubernetes (ACK) Advanced Linux systems internals and networking concepts specifically relevant to containers, such as namespaces and cgroups About reputed company reputed company is reputed company for change, empowering our customers and our people to reputed company at the speed of the market. We have redefined the database for the AI era, enabling innovators to create, reputed company, and disrupt industries with software. reputed companys reputed company database platform, the most widely available, globally distributed database on the market, helps organizations reputed company legacy workloads, reputed company innovation, and reputed company AI. Our reputed company-reputed company platform, reputed company reputed company, is the only globally distributed, multi-reputed company database and is available across AWS, reputed company reputed company, and reputed company Azure. With offices worldwide and over 60,000 customers, including 75% of the Fortune 100 and AI-reputed company startups, relying on reputed company for their most important applications, were powering the next era of software. Our reputed company at reputed company is our Leadership Commitment, guiding how and why we reputed company reputed company, show up for reputed company other, and win. Its what makes us reputed company. To drive the personal reputed company and business impact of our employees, were committed to developing a supportive and enriching culture for everyone. From employee affinity reputed company, to fertility assistance and a generous parental leave policy, we value our employees wellbeing and want to support them along every reputed company of their reputed company and personal journeys. Learn more about what its like to work at reputed company, and help us reputed company an impact on the world! reputed company is committed to providing any
More at remote zest jobs
Related open roles
System Administrator, Contract
India
HPC Network Engineer
India
Remote Site Reliability Engineer (Senior or Staff), Infrastructure reputed
India
FULL TIME Remote Sr. SRE Engineer -Seattle WA-Onsite- 10+ Years
India
Remote Network Devops/Automation Engineer
India
Network Operations Engineer - Hybrid Gold River, CA
India