Source description
About the role
Role: SRE & Deployments Engineer (SDE2) Function: Site Reliability Engineering / DevOps / Platform Engineering Location: Bengaluru Type: Full-time Industry: Information Technology & Services, Computer Software, Fintech About Company A Bengaluru-based enterprise tech startup founded in 2016. The company powers India's digital transformation through paperless, regulatory-compliant infrastructure. Its services include Aadhaar-enabled e-signing, e-KYC, eNACH, document automation, and consent management. It serves 1,500+ enterprises and 100 million individual users. The company is RBI-authorized as a payment aggregator and certified across security and compliance frameworks. A lean team of 90 engineers operates at the intersection of speed and stringent regulation. Position Overview This is a founding role in the company's Platform Pod — a newly formed team mandated to own the infrastructure every product team depends on. You will build and operate on-call, alerting, and incident response across pods, own SLOs for compliance-critical services, and own the end-to-end deployment pipeline across cloud and client-managed environments. You will also build the access-governance tooling that structurally eliminates the need for standing production access across the engineering org. Role & Responsibilities Build and operate on-call rotations, alerting pipelines, and incident response workflows across product pods; drive post-incident reviews and structural fixes to reduce repeat pages Define, instrument, and enforce SLOs for compliance-critical services; own error budgets and escalation paths Own the end-to-end deployment pipeline: reproducible packaging, artifact signing, SBOM generation, and install/upgrade automation across cloud and restricted-access environments Build remote observability and diagnostics tooling for deployments where direct environment access is limited or forbidden Build and maintain access-governance infrastructure: audited break-glass flows, session recording, and deploy-only paths that eliminate standing production access Operate and migrate core shared services with full runbook coverage, capacity planning, and DR validation Maintain and improve IaC, environment parity, and deployment automation across cloud and on-prem/hybrid environments Must Have Criteria 3–5 years in a DevOps or SRE role with real uptime accountability — carried a pager and made structural changes to reduce alert volume or MTTR Hands-on Linux administration and container operations in production environments Infrastructure-as-Code experience managing cloud resources at production scale Proficiency with at least one major cloud provider at an infrastructure operations level Scripting fluency in at least one language for automation, tooling, and diagnostic workflows Demonstrated experience with observability stacks — metrics, logging, and alerting pipelines in production Nice to Have Experience deploying and operating software in on-prem, hybrid, or customer-managed environments where direct access is restricted Exposure to regulated environments (banking, NBFC, or fintech) with awareness of audit, access-control, and compliance constraints Familiarity with artifact signing, SBOM generation, or supply-chain security practices Experience building or operating break-glass access systems, session recording, or privileged access management tooling Prior work in a platform or internal developer experience team serving engineering orgs of 30+ engineers What We Offer Founding-team scope: greenfield mandate to define how the Platform Pod operates — architectural decisions with lasting impact Direct ownership over systems used by 100M+ users and 1,500+ enterprises across India's payments and identity stack Flat, rotation-friendly engineering org (40 engineers) with real ownership and no bureaucratic gatekeeping High craft standards, low tolerance for toil — fast-moving team operating inside a regulated, high-consequence domain
More at recrew ai