Padmi

Intern Data Engineer

MumbaiPosted 3 months ago
Software engineeringInternFull Time; Regular
Apply at bluCognition

Opens the source posting on shine.com

Source description

About the role

View original

As a Data Engineer / Analytics Engineer at bluCognition, your role will involve: - Working with large structured datasets using SQL and PySpark. - Building, maintaining, and optimizing ETL/data processing pipelines. - Assisting in business/entity matching logic and fuzzy matching implementations. - Creating and validating analytical datasets for model development and reporting. - Performing data cleaning, transformation, aggregation, and quality checks. - Writing productive SQL queries using joins, CTEs, window functions, and aggregations. - Supporting feature engineering for ML/risk modeling use cases. - Working on incremental data processing and monthly/daily refresh strategies. - Analyzing data discrepancies, debugging pipeline failures, and improving reliability. - Collaborating with analytics, data science, and engineering teams. - Participating in testing, deployment, and code review activities. To help us level up, you ideally have: - A background in Computer Science, Data Science, Statistics, Mathematics, or a related field. - Solid SQL knowledge joins, CTEs, aggregations, CASE statements, and window functions. - Basic understanding of Python and familiarity with PySpark or distributed data processing concepts. - Understanding of relational databases. About bluCognition: bluCognition is an AI/ML based start-up specializing in risk analytics, data conversion, and data enrichment capabilities. Founded in 2017 by senior professionals from the financial services industry, the company is headquartered in the US with the delivery center based in Pune. We leverage the latest technology stack in AI, ML, and NLP combined with decades of experience in risk management to cater to some of the biggest names in the financial services industry. Join us in our exciting growth phase and be a part of our motivated and analytical team. As a Data Engineer / Analytics Engineer at bluCognition, your role will involve: - Working with large structured datasets using SQL and PySpark. - Building, maintaining, and optimizing ETL/data processing pipelines. - Assisting in business/entity matching logic and fuzzy matching implementations. - Creating and validating analytical datasets for model development and reporting. - Performing data cleaning, transformation, aggregation, and quality checks. - Writing productive SQL queries using joins, CTEs, window functions, and aggregations. - Supporting feature engineering for ML/risk modeling use cases. - Working on incremental data processing and monthly/daily refresh strategies. - Analyzing data discrepancies, debugging pipeline failures, and improving reliability. - Collaborating with analytics, data science, and engineering teams. - Participating in testing, deployment, and code review activities. To help us level up, you ideally have: - A background in Computer Science, Data Science, Statistics, Mathematics, or a related field. - Solid SQL knowledge joins, CTEs, aggregations, CASE statements, and window functions. - Basic understanding of Python and familiarity with PySpark or distributed data processing concepts. - Understanding of relational databases. About bluCognition: bluCognition is an AI/ML based start-up specializing in risk analytics, data conversion, and data enrichment capabilities. Founded in 2017 by senior professionals from the financial services industry, the company is headquartered in the US with the delivery center based in Pune. We leverage the latest technology stack in AI, ML, and NLP combined with decades of experience in risk management to cater to some of the biggest names in the financial services industry. Join us in our exciting growth phase and be a part of our motivated and analytical team.

One address, no account. We’ll tell you when matching roles go live.

More at bluCognition

Related open roles

View all roles