Source description
About the role
Responsibilities: Pipeline development: Design, build, and maintain end-to-end data pipelines for ingestion, transformation, and loading using Python and Databricks.SQL and data modelling: Write efficient SQL queries and scripts for data extraction, transformation, and validation across relational and cloud data stores.Azure data services: Build and manage pipelines leveraging Azure services including Azure Data Factory, Azure Blob Storage, and CosmosDB where applicable.Ad hoc data engineering: Respond to ad hoc data requests and pipeline changes raised by the Project Lead within agreed hour estimates.Data quality and validation: Implement checks, reconciliation steps, and logging to ensure data integrity and traceability across all pipelines.Orchestration and scheduling: Configure and manage job orchestration and scheduling on Databricks and Azure-native services.Documentation: Maintain clear technical documentation, pipeline design notes, data dictionaries, and runbooks to support knowledge transfer and continuity.Collaboration: Work closely with the Data Scientist, API Developer, and Cloud Engineer within the pod; coordinate with client stakeholders via the Project Lead only. Requirements: Education: Graduates only (please ignore pursuing/drop out/12th or 10th pass). .
More at Netscribes