Source description
About the role
Job Title: AWS Data Engineer Experience: 5 to 9 Years Location: Remote Shift Timing: 11 AM 8 PM Employment Type: Full-time Mandate Skillsets : Python, Pyspark, AWS Lambda, Glue & S3, Kinesis Streaming, Terraform SQL, CI/CD Pipeline & Iceburg Key Responsibilities: Design, implement, and maintain ETL pipelines on AWS using Step Functions, Glue, Lambda, and S3, aligned with Iceberg-based Lakehouse patterns. Build and optimize SQL transformations for Athena and Iceberg tables, including implementation of star and snowflake schemas from YAML table definitions and SQL transformations. Develop and maintain data models (fact and dimension tables) that power analytics and reporting in the data lake. Operate data ingestion pipelines and manage Lake Formation permissions to ensure secure and governed data access. Implement data quality checks and alerting using the in-repository data quality framework and AWS services. Provision, maintain, and evolve infrastructure using Terraform /AWS CDK as the Infrastructureas-Code framework. Contribute to CI/CD pipelines, including testing, linting, and deployments. Ensure data security and compliance through KMS encryption, IAM, Lake Formation controls, and audit logging. Monitor pipelines and infrastructure using CloudWatch and event-driven alerts. Collaborate effectively with cross-functional stakeholders using Atlassian project management tools, demonstrating strong stakeholder management and clear communication throughout the delivery lifecycle. Leverage AI-assisted development tools to accelerate troubleshooting, code reviews, and pipeline optimization while maintaining high quality and security standards. Required Skills: A bachelors or masters degree in a technical field. 5+ years of experience in data engineering or a similar role. Strong AWS data engineering experience, including S3,Glue, Athena, Lambda, Step Functions, Lake Formation, KMS, and CloudWatch. Strong experience in building real time data streaming pipelines. Proficiency in Python. Experience with data lake architecture, dbt, and Apache Iceberg (or similar table formats). Familiarity with CI/CD and DevOps practices. Comfort using AI tools responsibly within software development workflows. About IGT Atain: Atain is an enterprise orchestration partner that aligns people, processes, and platforms with AI to redefine work and deliver assured outcomes. It is enabled through the SMART Frameworkdriving superior metrics with AI-redefined work, the TechBud platform, a GenAI foundation built on zero-pilot architecture and ISO 42001 standards, a robust partner ecosystem, and deep domain expertise. Atain's Unified Hub continuously senses, predicts, and re-orchestrates work across people, processes, and platforms to deliver assured outcomes. The centers of skill and scale combine deep expertise with AI to deliver assured outcomes. For more information, visit www.atain.com Atain is ISO 27001:2013, CMMI SVC Level 5 and ISAE-3402 compliant for IT, and COPC Certified v6.0, ISO 27001:2013 and PCI DSS 3.2 certified for BPO processes. The organization follows Six Sigma rigor for process improvements. It is our policy to provide equal employment opportunities to all individuals based on job-related qualifications and ability to perform a job, without regard to age, gender, gender identity, sexual orientation, race, color, religion, creed, national origin, disability, genetic information, veteran status, citizenship or marital status, and to maintain a non-discriminatory environment free from intimidation, harassment or bias based upon these grounds.
More at ATAIN