Source description
About the role
As a Python Developer focusing on Data Science, you will be responsible for the following key tasks: - Design, develop, and maintain ETL pipelines for batch and streaming data. - Build robust data extraction, transformation, and loading workflows using Python. - Implement data validation, cleansing, normalization, and enrichment logic. - Manage incremental loads, backfills, and data quality checks. - Develop and maintain Druid ingestion specs for both batch and streaming data. - Optimize Druid data models, rollups, aggregations, and segment design. - Work with various Druid components such as Broker, Historical, MiddleManager, Indexer, and Coordinator. - Troubleshoot performance issues, segment merges, compaction, and schema evolution. Qualifications required for this role include: - 6-8 years of experience in Python development for Data Science. - Strong ETL development experience. - Working knowledge of Apache Druid. - Hands-on experience with Python, ETL frameworks, REST services, SQL, Druid ingestion specs, and modern cloud environments. The company is a top MNC client located in Chennai, TN, looking for a skilled Python Developer who can handle scalable data ingestion pipelines, transform large datasets, optimize analytics queries, and integrate services with Druid for real-time and batch analytics. As a Python Developer focusing on Data Science, you will be responsible for the following key tasks: - Design, develop, and maintain ETL pipelines for batch and streaming data. - Build robust data extraction, transformation, and loading workflows using Python. - Implement data validation, cleansing, normalization, and enrichment logic. - Manage incremental loads, backfills, and data quality checks. - Develop and maintain Druid ingestion specs for both batch and streaming data. - Optimize Druid data models, rollups, aggregations, and segment design. - Work with various Druid components such as Broker, Historical, MiddleManager, Indexer, and Coordinator. - Troubleshoot performance issues, segment merges, compaction, and schema evolution. Qualifications required for this role include: - 6-8 years of experience in Python development for Data Science. - Strong ETL development experience. - Working knowledge of Apache Druid. - Hands-on experience with Python, ETL frameworks, REST services, SQL, Druid ingestion specs, and modern cloud environments. The company is a top MNC client located in Chennai, TN, looking for a skilled Python Developer who can handle scalable data ingestion pipelines, transform large datasets, optimize analytics queries, and integrate services with Druid for real-time and batch analytics.
More at VMC Soft Technologies