Source description
About the role
Exp:5+ Years Job Location: Hyderabad(WFO) JD: Youll work closely with our Platform Architect with an initial focus on Data replication and CDC for time-sliced access to data Implementing data pipelines for different consumers Designing and implementing our Lakehouse (Iceberg+ Trino) Provisioning data specific features (e.g. databases, feature stores, model registries, ) to support AI/ML initiatives Enforcing data quality and lineage across batch & streaming Scaling across deployments (SaaS, hybrid, federated) with interoperability & integration in mind Result offering data stores and access APIs Qualifications 5+ years in data infrastructure, platform engineering, or related backend roles. Deep knowledge of o Data replication & CDC o Distributed processing (Spark/Trino) o Lakehouse/warehouse design (Iceberg/Delta + Trino) o Efficient storage design (partitioning, Z-ordering, compaction) o Schema & data contract governance o Document stores (cosmos DB, MongoDB) Proficient in Python, solid engineering fundamentals (testing, modularity, performance) Familiarity with infrastructure as code (Bicep) Familiarity with the Modern Data Stack Observability mindset: metrics, tracing, logging for data systems .
More at Talent Smart