Source description
About the role
Excellent knowledge on AWS Cloud platform as a whole and the integrations among different Cloud components and using one or more of AWS data and analytics services in combination with Spark, EMR, DynamoDB, RedShift, Athena, Lambda, Glue, Snowflake Design and build production data pipelines from ingestion to consumption within a big data architecture, using any programming language like Java, Python, Scala Design and implement data engineering, ingestion and curation functions on AWS cloud using AWS native or custom programming Perform detail assessments of current state data platforms and create an appropriate transition path to AWS cloud. Should have good experience in database, data warehouse concepts, SCD1, SDC2, SQLs Design, implement and support an analytical data infrastructure providing ad-hoc access to large datasets and computing power Interface with other technology teams to extract, transform, and load data from a wide variety of data sources using SQL and AWS big data technologies. Creation and support of real-time data pipelines built on AWS technologies including Glue, Redshift/Spectrum, Kinesis, EMR and Athena
EC2, Cloudfront, VPC, C++, Problem Solving
More at rhsandbox