Padmi

Senior Data Engineer (4-8 YOE) | Immediate Joiners (Mumbai)

MumbaiPosted 2 months ago
Infrastructure And DatabasesSeniorFull Time; Regular
Apply at YMinds.AI

Opens the source posting on shine.com

Source description

About the role

View original

About the Role Our client is seeking a highly experienced Senior Data Engineer (4-8 years) to join their Data Platform team. This role focuses on designing, building, and optimizing large-scale data infrastructure, pipelines, and cloud-based data platforms. The ideal candidate will work across the full data lifecycleenabling seamless data integration, scalable processing, and advanced analyticswhile supporting cross-functional teams and ensuring high-performance, reliable data systems for business-critical applications. Location : Gurugram and Chennai Key Responsibilities - Design, deploy, configure, and manage multi-node big data clusters across Dev, Test, and Production environments- Build and maintain scalable data pipelines, ETL/ELT workflows, and data lake architectures- Integrate data from multiple sources ensuring data quality, reliability, and accessibility- Develop automation scripts using Python and Shell scripting to streamline infrastructure and operations- Work with distributed data processing frameworks such as Apache Spark and Hadoop ecosystem- Design and maintain data models (dimensional modeling, star/snowflake schemas) for analytics and reporting- Develop and support BI solutions, dashboards, and KPIs for business insights- Optimize platform performance (scalability, latency, availability) and resolve system issues- Monitor, debug, and maintain production data systems and pipelines- Build tools and frameworks for ETL monitoring, validation, and troubleshooting- Collaborate with cross-functional teams including developers, analysts, and business stakeholders- Support production environments and on-call operations- Perform code reviews, data validation, and QA checks- Continuously evaluate and adopt current tools and technologies to improve data platform capabilities Required Skills & Experience - 4-8 years of experience in Data Engineering / Big Data environments- Strong expertise in Data Lake technologies:- Apache Spark, Hadoop, Yarn, Distributed File Systems- Experience with Cloud Platforms:- GCP (BigQuery, Dataflow, Cloud Storage) or AWS (Redshift, EMR)- Robust proficiency in SQL (Vertica, Dremio, or similar big data SQL engines)- Hands-on experience with Python and Shell scripting- Experience building ETL pipelines and OLAP systems- Experience with workflow orchestration tools (e.g., Apache Airflow)- Strong understanding of data warehousing concepts- Experience in data modeling for analytics and reporting- Strong debugging, monitoring, and troubleshooting skills- Familiarity with version control systems and relational databases- Experience working in Linux-based environments- Bachelors/Masters degree in Computer Science or related field Nice-to-Have Skills - Experience with Visualization tools (Apache Superset, Tableau)- Exposure to AWS EMR and advanced cloud services- Strong understanding of data structures and analytics workflows- Experience working in Agile environments- Experience with on-call production support- Strong communication and system design skills About YMinds.AI YMinds.AI is a talent and consulting partner connecting top-tier tech professionals with high-growth organizations. We specialize in building high-performing engineering teams by aligning the right talent with the right opportunities, ensuring long-term success for both clients and candidates. Keywords Senior Data Engineer, Big Data, Data Lake, Apache Spark, Hadoop, Yarn, GCP, AWS, SQL, Python, ETL, OLAP, Data Modeling, BI, Data Pipelines, Airflow, BigQuery, Redshift Hashtags #DataEngineering #SeniorDataEngineer #BigData #ApacheSpark #Hadoop #GCP #AWS #Python #SQL #Hiring #ChennaiJobs #TechJobs #DataJobs About the Role Our client is seeking a highly experienced Senior Data Engineer (4-8 years) to join their Data Platform team. This role focuses on designing, building, and optimizing large-scale data infrastructure, pipelines, and cloud-based data platforms. The ideal candidate will work across the full data lifecycleenabling seamless data integration, scalable processing, and advanced analyticswhile supporting cross-functional teams and ensuring high-performance, reliable data systems for business-critical applications. Location : Gurugram and Chennai Key Responsibilities - Design, deploy, configure, and manage multi-node big data clusters across Dev, Test, and Production environments- Build and maintain scalable data pipelines, ETL/ELT workflows, and data lake architectures- Integrate data from multiple sources ensuring data quality, reliability, and accessibility- Develop automation scripts using Python and Shell scripting to streamline infrastructure and operations- Work with distributed data processing frameworks such as Apache Spark and Hadoop ecosystem- Design and maintain data models (dimensional modeling, star/snowflake schemas) for analytics and reporting- Develop and support BI solutions, dashboards, and KPIs for business insights- Optimi

One address, no account. We’ll tell you when matching roles go live.

More at YMinds.AI

Related open roles

View all roles