Padmi

SRE (Blockchain, Web3 & Ai Native)

Delhi NCRPosted 3 months ago
Software engineeringSeniorFull Time; Regular
Apply at InfraSingularity

Opens the source posting on shine.com

Source description

About the role

View original

As a Senior Site Reliability Engineer (SRE) at InfraSingularity, you will play a crucial role in managing our multi-cloud blockchain infrastructure and validator node operations. Your expertise in infrastructure automation and system reliability will be instrumental in ensuring the high performance, availability, and resilience of various L1/L2 blockchain protocols. If you are passionate about emerging Web3 technologies, this is the perfect opportunity for you to make a significant impact. Responsibilities: - Own and operate validator nodes across multiple blockchain networks, ensuring uptime, security, and cost-efficiency. - Architect, deploy, and maintain infrastructure on AWS, GCP, and bare-metal for protocol scalability and performance. - Implement Kubernetes-native tooling (Helm, FluxCD, Prometheus, Thanos) to manage deployments and observability. - Collaborate with our Protocol R&D team to onboard new blockchains and participate in testnets, mainnets, and governance. - Ensure secure infrastructure with best-in-class secrets management (HashiCorp Vault, KMS) and incident response protocols. - Contribute to a robust monitoring and alerting stack to detect anomalies, performance drops, or protocol-level issues. - Act as a bridge between software, protocol, and product teams to communicate infra constraints or deployment risks clearly. - Continuously improve deployment pipelines using Terraform, Terragrunt, GitOps practices. - Participate in on-call rotations and incident retrospectives, driving post-mortem analysis and long-term fixes. Our Stack: - Cloud & Infra: AWS, GCP, bare-metal - Containerization: Kubernetes, Helm, FluxCD - IaC: Terraform, Terragrunt - Monitoring: Prometheus, Thanos, Grafana, Loki - Secrets & Security: HashiCorp Vault, AWS KMS - Languages: Go, Bash, Python, Types - Blockchain: Ethereum, Polygon, Cosmos, Solana, Foundry, OpenZeppelin What You Bring: - 4+ years of experience in SRE/DevOps/Infra rolesideally within FinTech, Cloud, or high-reliability environments. - Proven expertise managing Kubernetes with strong hands-on experience in Terraform, Helm, GitOps workflows. - Deep understanding of system reliability, incident management, fault tolerance, and monitoring best practices. - Proficiency with Prometheus and PromQL for custom dashboards, metrics, and alerting. - Experience operating secure infrastructure and implementing SOC2/ISO27001-aligned practices. - Solid scripting skills in Bash, Python, or Go. - Clear and confident communicator capable of interfacing with both technical and non-technical stakeholders. Nice-to-Have: - First-hand experience in Web3/blockchain/crypto environments. - Understanding of staking, validator economics, slashing conditions, or L1/L2 governance mechanisms. - Exposure to smart contract deployments or working with Solidity, Foundry, or similar toolchains. - Experience with compliance-heavy or security-certified environments (SOC2, ISO 27001, HIPAA). If you join InfraSingularity, you will have the opportunity to work at the forefront of Web3 infrastructure and validator technology, collaborate with a dynamic team that values ownership and performance, and gain exposure to some of the most exciting blockchain ecosystems globally. As a Senior Site Reliability Engineer (SRE) at InfraSingularity, you will play a crucial role in managing our multi-cloud blockchain infrastructure and validator node operations. Your expertise in infrastructure automation and system reliability will be instrumental in ensuring the high performance, availability, and resilience of various L1/L2 blockchain protocols. If you are passionate about emerging Web3 technologies, this is the perfect opportunity for you to make a significant impact. Responsibilities: - Own and operate validator nodes across multiple blockchain networks, ensuring uptime, security, and cost-efficiency. - Architect, deploy, and maintain infrastructure on AWS, GCP, and bare-metal for protocol scalability and performance. - Implement Kubernetes-native tooling (Helm, FluxCD, Prometheus, Thanos) to manage deployments and observability. - Collaborate with our Protocol R&D team to onboard new blockchains and participate in testnets, mainnets, and governance. - Ensure secure infrastructure with best-in-class secrets management (HashiCorp Vault, KMS) and incident response protocols. - Contribute to a robust monitoring and alerting stack to detect anomalies, performance drops, or protocol-level issues. - Act as a bridge between software, protocol, and product teams to communicate infra constraints or deployment risks clearly. - Continuously improve deployment pipelines using Terraform, Terragrunt, GitOps practices. - Participate in on-call rotations and incident retrospectives, driving post-mortem analysis and long-term fixes. Our Stack: - Cloud & Infra: AWS, GCP, bare-metal - Containerization: Kubernetes, Helm, FluxCD - IaC: Terraform, Terragrunt - Monitoring: Prometheus, Thanos, Gr

One address, no account. We’ll tell you when matching roles go live.

More at InfraSingularity

Related open roles

View all roles