Source description
About the role
Area(s) of responsibility Role Summary We are looking for a highly skilled Azure Data Engineer with strong hands on expertise in Azure Databricks, PySpark, ADF, and SQL. The candidate will be responsible for architecting, developing, and optimizing modern data platforms and analytical solutions on Azure. The role requires strong experience building enterprise-grade data pipelines, enabling ingestion from diverse sources, implementing complex transformations, and supporting scalable analytics initiatives. Key Responsibilities - Data Solution Design & Architecture Design end to end data engineering solutions using Azure Data Factory, Azure Databricks, PySpark, SQL, and other Azure-native services. Architect and implement scalable and secure Contemporary Data Warehouse (MDW) and Lakehouse solutions leveraging Azure Data Lake Storage and Databricks. Develop data models, integration patterns, and reusable frameworks aligned with best practices and enterprise architecture standards. Participate in requirement discussions, solution blueprinting, and technical feasibility assessments. - Data Pipeline Development Build and optimize robust, high-throughput ELT/ETL pipelines, enabling ingestion, transformation, and curation of structured, semi structured, and unstructured data. Integrate data from multiple on premise and cloud based systems, APIs, and third-party sources. Implement complex transformations using PySpark, ensuring performance efficiency and code modularity. Build orchestration workflows in ADF, including pipelines, triggers, linked services, integration runtimes, and parameterized datasets. - Databricks & PySpark Engineering Develop scalable transformation scripts using PySpark on Databricks, applying advanced optimizations like caching, partitioning, and Delta Lake capabilities. Implement Delta Lake featuresACID transactions, schema enforcement, schema evolution, and time travelacross the data lifecycle. Perform performance tuning, handling bottleneck .
More at Birlasoft