Source description
About the role
Data Engineer | Hadoop, Spark, PySpark, Python Experience: 10+ Years Location: Chennai / Bangalore Work Mode: Hybrid Job Description We are hiring an experienced Data Engineer with strong expertise in Hadoop, Apache Spark, PySpark, Python, Hive, Advanced SQL, and Unix/Linux . The ideal candidate will design, develop, and optimize scalable data pipelines and big data solutions while working with large-scale distributed data processing systems. Key Responsibilities Design and develop scalable ETL/ELT data pipelines using PySpark and Apache Spark. Build and optimize data processing workflows on the Hadoop ecosystem. Develop data ingestion and transformation frameworks using Python and Spark. Write optimized SQL queries for data extraction, transformation, and reporting. Develop and maintain Hive tables, partitioning, and performance tuning. Monitor, troubleshoot, and optimize production data pipelines. Perform data validation, quality checks, and root cause analysis. Collaborate with cross-functional teams in an Agile environment. Mandatory Skills Apache Hadoop Apache Spark & Spark Core PySpark Python Hive Advanced SQL Unix/Linux & Shell Scripting ETL/ELT Development Big Data Technologies Data Pipeline Development Preferred Qualifications BE/B.Tech, MCA, M.Tech, or equivalent. Strong analytical and problem-solving skills. Experience working in Agile/Scrum environments. Interested candidate share me your updated cv Email-it.hr5@evokehr.com Whatsapp-7226857259
More at Evoke HR