Source description
About the role
Role:Data Engineer Location:Remote Duration:6 month contract with extensions to 2 years must haves 1. SQL 2. Python 3. Spark 4. Hadoop 5. ETL 6. Unix 7. PySpark and Hive (they should have Hive if they have PySpark) 8. CI/CD – they use Git and Jenkins but open to other experience on this Job Description: Project Details: They’re modernizing data quality controls and regression tests for financial crimes data sets. DevOps engineer family and specializing in automation but not typical automation. Automating data validation scripts on big data. The ETL pipeline is already built. They have the template with 3 instances of it so this should be the 4th instance. Only expectation is building and comparing it against what the spark programs are doing. Preferred qualifications: 1. Database modeling 2. Building schemas Communication skills are good and team oriented mindset Thanks & Regards, PremTarun P IT-Recruiter Staffing | Thoughtwave Software and Solutions Desk: 6306182581 | Email:tarun@wavsols.com
More at Thoughtwave Software and Solutions