Padmi

Lead GCP Data Engineer

HyderabadPosted 1 month ago
Software engineeringSeniorFull Time; Regular
Apply at People Prime Worldwide

Opens the source posting on shine.com

Source description

About the role

View original

Educational QualificationBachelor's degree in Computer Science, or a related field, or equivalent experienceKey ResponsibilitiesDesign, develop, test, and maintain scalable ETL data pipelines using PythonArchitect enterprise solutions with technologies including Kafka, GCP services, GKE, Load Balancers, APIGEE, DBT, LLMs, data redaction, and DLP solutionsBuild and manage data workflows on Google Cloud Platform using Dataflow, Cloud Functions, BigQuery, Cloud Composer, Cloud Storage, IAM, and Cloud RunImplement data ingestion, transformation, and cleansing to ensure high-quality data deliveryDevelop and enforce data quality checks, validation rules, and monitoring processesCollaborate with data scientists, analysts, and engineering teams to deliver efficient data solutionsManage version control using GitHub and participate in CI/CD pipeline deploymentsWrite complex SQL queries for data extraction and validation from relational databases (SQL Server, Oracle, PostgreSQL)Document pipeline designs, data flow diagrams, and operational support proceduresRequired Skills & CompetenciesStrong hands-on experience in Python for backend or data engineering projectsExpertise in GCP services: Dataflow, BigQuery, Cloud Functions, Cloud Composer, Cloud Storage, Cloud Run, IAMExperience in data pipeline architecture, data integration, and transformationsProficiency in Apache Spark, Kafka, Redis/Bigtable, FastAPI, and AirflowStrong SQL skills with at least one enterprise database (SQL Server, Oracle, PostgreSQL)Experience in CI/CD practices and version control (GitHub)Good to Have / Optional SkillsExperience with Snowflake cloud data platformHands-on knowledge of Databricks for big data processing and analyticsFamiliarity with Azure Data Factory (ADF) and Azure data engineering tools Educational QualificationBachelor's degree in Computer Science, or a related field, or equivalent experienceKey ResponsibilitiesDesign, develop, test, and maintain scalable ETL data pipelines using PythonArchitect enterprise solutions with technologies including Kafka, GCP services, GKE, Load Balancers, APIGEE, DBT, LLMs, data redaction, and DLP solutionsBuild and manage data workflows on Google Cloud Platform using Dataflow, Cloud Functions, BigQuery, Cloud Composer, Cloud Storage, IAM, and Cloud RunImplement data ingestion, transformation, and cleansing to ensure high-quality data deliveryDevelop and enforce data quality checks, validation rules, and monitoring processesCollaborate with data scientists, analysts, and engineering teams to deliver efficient data solutionsManage version control using GitHub and participate in CI/CD pipeline deploymentsWrite complex SQL queries for data extraction and validation from relational databases (SQL Server, Oracle, PostgreSQL)Document pipeline designs, data flow diagrams, and operational support proceduresRequired Skills & CompetenciesStrong hands-on experience in Python for backend or data engineering projectsExpertise in GCP services: Dataflow, BigQuery, Cloud Functions, Cloud Composer, Cloud Storage, Cloud Run, IAMExperience in data pipeline architecture, data integration, and transformationsProficiency in Apache Spark, Kafka, Redis/Bigtable, FastAPI, and AirflowStrong SQL skills with at least one enterprise database (SQL Server, Oracle, PostgreSQL)Experience in CI/CD practices and version control (GitHub)Good to Have / Optional SkillsExperience with Snowflake cloud data platformHands-on knowledge of Databricks for big data processing and analyticsFamiliarity with Azure Data Factory (ADF) and Azure data engineering tools

One address, no account. We’ll tell you when matching roles go live.

More at People Prime Worldwide

Related open roles

View all roles