Source description
About the role
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. Must have skills : PySpark Good to have skills : NA Minimum 3 Year(s) Of Experience Is Required Educational Qualification : 15 years full time education Summary As a Data Engineer, a typical day involves designing, developing, and maintaining comprehensive data solutions that support the generation, collection, and processing of data. The role includes creating efficient data pipelines to facilitate smooth data flow and ensuring the integrity and quality of data throughout its lifecycle. Additionally, the position requires implementing processes to extract, transform, and load data, enabling seamless migration and deployment across various systems. This role demands continuous collaboration with different teams to optimize data handling and support organizational data needs effectively. Roles & Responsibilities - Expected to perform independently and become an SME. - Required active participation/contribution in team discussions. - Contribute in providing solutions to work related problems. - Collaborate with cross-functional teams to understand data requirements and deliver scalable data solutions. - Monitor and troubleshoot data pipelines to ensure reliability and performance. - Document data processes and workflows to maintain clarity and support knowledge sharing. - Assist junior team members in understanding project requirements and technical challenges. Professional & Technical Skills: - Must To Have Skills: Proficiency in PySpark. - Experience in building and optimizing data pipelines using distributed computing frameworks. - Robust knowledge of data processing concepts including batch and real-time data handling. - Familiarity with data storage solutions and data modeling techniques. - Ability to write efficient, maintainable code for data transformation and integration. - Understanding of data quality assurance and validation methods. Additional Information - The candidate should have minimum 3 years of experience in PySpark. - This position is based at our Kolkata office. - A 15 years full time education is required. , 15 years full time education Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. Must have skills : PySpark Good to have skills : NA Minimum 3 Year(s) Of Experience Is Required Educational Qualification : 15 years full time education Summary As a Data Engineer, a typical day involves designing, developing, and maintaining comprehensive data solutions that support the generation, collection, and processing of data. The role includes creating efficient data pipelines to facilitate smooth data flow and ensuring the integrity and quality of data throughout its lifecycle. Additionally, the position requires implementing processes to extract, transform, and load data, enabling seamless migration and deployment across various systems. This role demands continuous collaboration with different teams to optimize data handling and support organizational data needs effectively. Roles & Responsibilities - Expected to perform independently and become an SME. - Required active participation/contribution in team discussions. - Contribute in providing solutions to work related problems. - Collaborate with cross-functional teams to understand data requirements and deliver scalable data solutions. - Monitor and troubleshoot data pipelines to ensure reliability and performance. - Document data processes and workflows to maintain clarity and support knowledge sharing. - Assist junior team members in understanding project requirements and technical challenges. Professional & Technical Skills: - Must To Have Skills: Proficiency in PySpark. - Experience in building and optimizing data pipelines using distributed computing frameworks. - Robust knowledge of data processing concepts including batch and real-time data handling. - Familiarity with data storage solutions and data modeling techniques. - Ability to write efficient, maintainable code for data transformation and integration. - Understanding of data quality assurance and validation methods. Additional Information - The candidate should have minimum 3 years of experience in PySpark. - This position is based at our Kolkata office. - A 15 years full time education is required. , 15 years full time education
More at Accenture