Padmi

Data Engineer - Python/Spark

IndiaPosted 3 months ago
Software engineeringMid-levelFull Time; Regular
Apply at Okda Solutions

Opens the source posting on shine.com

Source description

About the role

View original

You are a highly skilled Data Engineer with expertise in Procurement domain and Master Data Management (MDM). Your role involves building scalable data platforms, enabling procurement analytics, and ensuring high-quality master data across enterprise systems. You will work at the intersection of data engineering, governance, and business operations. - Design, build, and maintain scalable data pipelines for procurement and supply chain data - Develop and optimize ETL/ELT pipelines for large and complex datasets - Build and manage data lakes and data warehouses - Ensure high performance, reliability, and scalability of data systems - Develop and maintain MDM frameworks for supplier, vendor, and material master data - Ensure data quality, consistency, and standardization - Implement data cleansing, validation, and enrichment processes - Define and enforce data governance policies - Collaborate with Procurement, Finance, and Supply Chain teams - Manage supplier master, spend, contract, and catalog data - Enable analytics such as spend analysis and vendor performance - Integrate data from ERP systems and external sources - Establish and monitor data quality KPIs and SLAs - Resolve data inconsistencies and duplication issues - Support audits and compliance requirements - Work with data scientists, analysts, and product teams - Build self-service data platforms - Communicate complex data concepts to stakeholders You should have strong proficiency in Python and SQL, experience with ETL tools (Airflow, dbt), hands-on experience with data warehouses (Snowflake, BigQuery, Redshift), familiarity with Delta Lake / Iceberg, and a strong understanding of data modeling and governance. Experience with Apache Spark, Kafka, Hadoop, and MDM frameworks is essential. Preferred skills include exposure to SAP (MM), Ariba, Coupa, Ivalua, GEP, Zycus, understanding of procurement lifecycle, experience with AWS, Azure, or GCP, and experience in global environments. Your behavioral skills should include strong analytical and problem-solving skills, excellent communication skills, ability to translate business needs into data solutions, and a collaborative mindset. You are a highly skilled Data Engineer with expertise in Procurement domain and Master Data Management (MDM). Your role involves building scalable data platforms, enabling procurement analytics, and ensuring high-quality master data across enterprise systems. You will work at the intersection of data engineering, governance, and business operations. - Design, build, and maintain scalable data pipelines for procurement and supply chain data - Develop and optimize ETL/ELT pipelines for large and complex datasets - Build and manage data lakes and data warehouses - Ensure high performance, reliability, and scalability of data systems - Develop and maintain MDM frameworks for supplier, vendor, and material master data - Ensure data quality, consistency, and standardization - Implement data cleansing, validation, and enrichment processes - Define and enforce data governance policies - Collaborate with Procurement, Finance, and Supply Chain teams - Manage supplier master, spend, contract, and catalog data - Enable analytics such as spend analysis and vendor performance - Integrate data from ERP systems and external sources - Establish and monitor data quality KPIs and SLAs - Resolve data inconsistencies and duplication issues - Support audits and compliance requirements - Work with data scientists, analysts, and product teams - Build self-service data platforms - Communicate complex data concepts to stakeholders You should have strong proficiency in Python and SQL, experience with ETL tools (Airflow, dbt), hands-on experience with data warehouses (Snowflake, BigQuery, Redshift), familiarity with Delta Lake / Iceberg, and a strong understanding of data modeling and governance. Experience with Apache Spark, Kafka, Hadoop, and MDM frameworks is essential. Preferred skills include exposure to SAP (MM), Ariba, Coupa, Ivalua, GEP, Zycus, understanding of procurement lifecycle, experience with AWS, Azure, or GCP, and experience in global environments. Your behavioral skills should include strong analytical and problem-solving skills, excellent communication skills, ability to translate business needs into data solutions, and a collaborative mindset.

One address, no account. We’ll tell you when matching roles go live.