Padmi

Sr. Site Reliability Engineer  (Storage Platform)

United StatesPosted 1 month ago
Infrastructure And DatabasesUnspecified
Apply at 3B Staffing

Opens the source posting on 3bstaffing.com

Source description

About the role

View original

Job Title : Sr. Site Reliability Engineer (Storage Platform)

Client: AHEAD

VISA: US Citizens / GC

Salary: $66/Hr. on C2C

Job Type : Contract, with potential for Contract-to-hire – the client only wants to see candidates that are willing to convert to full-time employment for this role that do not require any type of sponsorship.

Worksite Requirement : Fully Remote

Industry : IT Services and IT Consulting

Interview Process : 2-3 rounds of video conference interviews

Job Summary

We are seeking a highly experienced Sr Site Reliability Engineer – Storage Platforms to design, implement, and support Software Defined Storage (SDS) and Kubernetes platforms in a private cloud environment. This role focuses on scalability, resilience, automation, and performance using Infrastructure-as-Code and GitOps practices.

This is a deeply technical role requiring expert-level understanding of Software Defined Storage, Kubernetes, and extensive working knowledge on Linux Operating systems . You will also collaborate with platform and SRE teams to maintain secure, performant, and multitenant-isolated services that serve high-throughput, mission-critical applications.

Key Responsibilities

Design, implement, and operate large-scale Software Defined Storage architectures across private and public cloud regions within ITIL methodology.

Deploy and support enterprise storage platforms (Pure Storage, HPE, NetApp) and SDS solutions (Ceph, Longhorn).

Build self-service storage workflows for Kubernetes CSI and OpenStack consumers (VM and Baremetal).

Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation.

Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback.

Build observability, alerting, and auto-remediation using GitOps and tools such as Prometheus, Loki, and Grafana.

Architect and maintain high availability, disaster recovery, and scale-out infrastructure.

Develop and review high-level and low-level design documents for storage infrastructure

Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems.

Participate in on-call rotations, incident response, and root cause analysis.

Collaborate globally on change management, documentation, and operational best practices.

Must Have

6+ years of experience managing enterprise storage and Kubernetes platforms on Linux.

Strong hands-on experience with SDS solutions (Ceph, Longhorn) and storage migrations from legacy systems.

Experience with block, file, and object storage, including Fibre Channel and IP-based protocols.

Experience with NVMe-oF or iSCSI fabrics.

Expert knowledge of Kubernetes and Linux systems (Ubuntu, RHEL/CentOS).

Proficiency with Infrastructure-as-Code (IaC) (Ansible, Terraform).

Strong scripting skills in Python and Bash (Golang (GO) a plus).

Strong working knowledge of Enterprise DNS and integrations with Kubernetes

Experience operating 24x7 mission-critical production environments.

Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack).

Strong written and verbal communication skills.

Proficiency with Git, CI/CD pipelines, and automated testing frameworks

Ability to write technical documentation and contribute to community wikis or knowledge bases.

Bachelor's degree in computer science or equivalent professional experience.

Nice to Have

OpenStack Cinder multi-backend administration.

Backup platforms (Rubrik).

Understanding of CIS/NIST security and infrastructure lifecycle management.

ITIL Foundation/advanced certifications in support of ITSM standard methodology.

Background in telco, edge cloud, or large enterprise environments.

CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security

Specialist (CKS) or Red Hat specialist in Ceph Storage Administrator (EX125) certifications.

Master's degree in computer science, IT, Engineering, or a related field preferred;

equivalent experience and relevant industry certifications will also be considered.

One address, no account. We’ll tell you when matching roles go live.

More at 3B Staffing

Related open roles

View all roles