Padmi

Sr. Data Engineer

MumbaiPosted 3 months ago
Infrastructure And DatabasesSeniorFull Time; Regular
Apply at Important Group

Opens the source posting on shine.com

Source description

About the role

View original

Role Overview: You are a highly experienced and motivated Senior Data Engineer with expertise in PySpark, ETL processes, and SQL. Your role will involve designing and building scalable data pipelines, data integration, and transformation workflows in distributed environments to enable data-driven decision-making across the organization. Key Responsibilities: - Design, develop, and maintain large-scale, distributed data processing systems using PySpark on big data platforms (Hadoop/Spark). - Build and automate ETL/ELT pipelines to extract data from various structured and unstructured sources. - Optimize and troubleshoot SQL queries for performance, scalability, and accuracy. - Collaborate with Data Architects, Data Scientists, and Analysts to deliver clean, structured, and reliable data for business use. - Implement best practices for data modeling, data quality, and data governance. - Monitor and enhance data workflows to ensure reliability and performance in production. - Contribute to technical discussions, architectural reviews, and code reviews. - Mentor junior data engineers and support their technical growth. Qualification Required: - 812 years of experience in Data Engineering roles. - Strong expertise in PySpark and Apache Spark for large-scale data processing. - Experience designing and building ETL pipelines in production environments. - Proficiency in writing complex and optimized SQL queries. - Familiarity with big data technologies (e.g., Hadoop, Hive, HDFS, Delta Lake). - Knowledge of data warehouse platforms like Snowflake, Redshift, or BigQuery is a plus. - Understanding of data architecture, data modeling, and data quality frameworks. - Experience with cloud platforms (AWS, Azure, or GCP) is preferred. - Strong problem-solving and debugging skills. - Excellent communication and collaboration skills. Additional Details (if any): Aligned Automation is a strategic service provider that partners with Fortune 500 leaders to digitize enterprise operations and enable business strategies. The company believes in creating positive, lasting change in the way clients work while advancing the global impact of their business solutions. The company culture is enriched by the values of Care, Courage, Curiosity, and Collaboration, supporting solutions that empower possibilities. Role Overview: You are a highly experienced and motivated Senior Data Engineer with expertise in PySpark, ETL processes, and SQL. Your role will involve designing and building scalable data pipelines, data integration, and transformation workflows in distributed environments to enable data-driven decision-making across the organization. Key Responsibilities: - Design, develop, and maintain large-scale, distributed data processing systems using PySpark on big data platforms (Hadoop/Spark). - Build and automate ETL/ELT pipelines to extract data from various structured and unstructured sources. - Optimize and troubleshoot SQL queries for performance, scalability, and accuracy. - Collaborate with Data Architects, Data Scientists, and Analysts to deliver clean, structured, and reliable data for business use. - Implement best practices for data modeling, data quality, and data governance. - Monitor and enhance data workflows to ensure reliability and performance in production. - Contribute to technical discussions, architectural reviews, and code reviews. - Mentor junior data engineers and support their technical growth. Qualification Required: - 812 years of experience in Data Engineering roles. - Strong expertise in PySpark and Apache Spark for large-scale data processing. - Experience designing and building ETL pipelines in production environments. - Proficiency in writing complex and optimized SQL queries. - Familiarity with big data technologies (e.g., Hadoop, Hive, HDFS, Delta Lake). - Knowledge of data warehouse platforms like Snowflake, Redshift, or BigQuery is a plus. - Understanding of data architecture, data modeling, and data quality frameworks. - Experience with cloud platforms (AWS, Azure, or GCP) is preferred. - Strong problem-solving and debugging skills. - Excellent communication and collaboration skills. Additional Details (if any): Aligned Automation is a strategic service provider that partners with Fortune 500 leaders to digitize enterprise operations and enable business strategies. The company believes in creating positive, lasting change in the way clients work while advancing the global impact of their business solutions. The company culture is enriched by the values of Care, Courage, Curiosity, and Collaboration, supporting solutions that empower possibilities.

One address, no account. We’ll tell you when matching roles go live.

More at Important Group

Related open roles

View all roles